One Frozen Simulator Is Not Enough: Simulator Collapse in Multi-Agent RL
DGX agentarXiv:2608.12253v1 Announce Type: cross Abstract: Multi-agent reinforcement learning for human-AI interaction typically relies on a single large language model to simulate user behavior. We show that