Breaking the Tokenizer Barrier: On-Policy Distillation across Model Families
DGX agentarXiv:2606.09456v1 Announce Type: new Abstract: On-Policy Distillation (OPD) has become a core technique in the post-training of Large Language Models (LLMs) for transferring knowledge from domain exp