epoch.ai

An FAQ on Reinforcement Learning Environments

dcre · 44 points · 9 comments · 3월 19일 · Open original

Comments

3 preview comments · loading full thread
nithisha22013월 25일

One thing missing from most RL environment discussions is observability during training. Single-agent envs are hard enough to debug, but multi-agent environments are a completely different challenge, reward curves tell you almost nothing about which agent failed or why cooperation broke down.

gizajob3월 21일

“An FAQ” really sets my grammar nerves jangling. “A FAQ” isn’t great either. Maybe “FAQ on Reinforcement…” or “FAQ about Reinforcement…”

idorosen3월 21일

ai slop