Metric-Gradient Projection for Stable Multi-Agent Policy Learning
DGX agentarXiv:2605.18809v1 Announce Type: cross Abstract: General-sum multi-agent learning is often governed by a stacked update field in which each agent's policy update changes the optimization landscape fa