Learning to Refine Hidden States for Reliable LLM Reasoning
DGX agentarXiv:2606.17524v2 Announce Type: replace Abstract: Large language models show strong reasoning ability, but their internal reasoning process can remain unstable in complex multi-step settings, where