Training Language Models to Cooperate with Inference-Time Controllers
DGX agentarXiv:2607.23771v1 Announce Type: new Abstract: Large language model (LLM) performance increasingly depends not only on the base model, but also on the inference-time controller used to organize reaso