LaTER: Efficient Test-Time Reasoning via Latent Exploration and Explicit Verification
DGX agentarXiv:2605.07315v1 Announce Type: new Abstract: Chain-of-thought (CoT) reasoning improves large language models (LLMs) on difficult tasks, but it also makes inference expensive because every intermedi