Entropy-Regularized Adjoint Matching for Offline Reinforcement Learning
DGX agentarXiv:2605.06156v2 Announce Type: replace-cross Abstract: Integrating expressive generative policies, such as flow-matching models, into offline reinforcement learning (RL) allows agents to capture co