Optimal Transport Q-Learning for Flow Policy Steering and Acceleration
arXiv:2607.06262v1 Announce Type: new Abstract: Diffusion and flow policies have recently demonstrated remarkable performance in robotic applications by accurately capturing multimodal robot trajector