epoch.ai

An FAQ on Reinforcement Learning Environments

dcre · 44 points · 9 comments · Mar 19 · Open original

Comments

3 preview comments · loading full thread
nithisha2201Mar 25

One thing missing from most RL environment discussions is observability during training. Single-agent envs are hard enough to debug, but multi-agent environments are a completely different challenge, reward curves tell you almost nothing about which agent failed or why cooperation broke down.

gizajobMar 21

“An FAQ” really sets my grammar nerves jangling. “A FAQ” isn’t great either. Maybe “FAQ on Reinforcement…” or “FAQ about Reinforcement…”

idorosenMar 21

ai slop