Monitoring the Internal Monologue: Probe Trajectories Reveal Reasoning Dynamics
DGX agentarXiv:2605.18549v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) introduce new opportunities for safety monitoring through their Chain of Thought (CoT) reasoning. However, CoT is not alwa