Reason Wide, Not Deep: Amortizing the Reasoning Premium into Distilled Skills
DGX agentarXiv:2608.07885v1 Announce Type: new Abstract: Reasoning modes of language models outperform their non-reasoning counterparts on multi-step agentic tasks, but pay a 3-6x premium in output tokens on e