When Context Returns: Toward Robust Internalization in On-Policy Distillation
DGX agentarXiv:2606.11627v1 Announce Type: cross Abstract: Recent work has shown that on-policy distillation can internalize privileged context, such as system prompts or task hints, into a student model so th