Coupled Variational Reinforcement Learning for Language Model General Reasoning
DGX agentarXiv:2512.12576v3 Announce Type: replace-cross Abstract: While reinforcement learning has achieved impressive progress in language model reasoning, it is constrained by the requirement for verifiable