Weak-Driven Learning: How Weak Agents make Strong Agents Stronger
DGX agentarXiv:2602.08222v2 Announce Type: replace Abstract: As post-training optimization becomes central to improving large language models, we observe a persistent saturation bottleneck: once models grow hi