Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents
DGX agentarXiv:2608.09555v1 Announce Type: new Abstract: External natural-language skills provide large language model (LLM) agents with reusable and editable guidance for solving complex tasks. Yet their effe