Precise: SDE-Consistent Stochastic Sampling for RL Post-Training of Flow-Matching Models
DGX agentarXiv:2605.23522v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become an effective way to improve prompt alignment and perceptual quality in diffusion and flow-matching generators.