epoch.ai

An FAQ on Reinforcement Learning Environments

dcre · 44 points · 9 comments · 19 thg 3 · Open original

Comments

3 preview comments · loading full thread
nithisha220125 thg 3

One thing missing from most RL environment discussions is observability during training. Single-agent envs are hard enough to debug, but multi-agent environments are a completely different challenge, reward curves tell you almost nothing about which agent failed or why cooperation broke down.

gizajob21 thg 3

“An FAQ” really sets my grammar nerves jangling. “A FAQ” isn’t great either. Maybe “FAQ on Reinforcement…” or “FAQ about Reinforcement…”

idorosen21 thg 3

ai slop