Training Small LLMs as Spatial Multi-Agent Policies
DGX agentarXiv:2608.01425v1 Announce Type: cross Abstract: Training LLM-based multi-agent systems with multi-agent reinforcement learning is rapidly gaining traction, and a parallel line of work argues that su