Masked Distillation: Internalizing the Chain-of-Thought in Language Models
DGX agentarXiv:2607.22629v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) produce long, explicit chains of intermediate steps before generating a final answer at inference time. These intermediate