RL-Native Distillation: Exploiting Scored Trajectories for Few-Step Image Generation
arXiv:2608.09226v1 Announce Type: cross Abstract: Efficient text-to-image generation requires both reinforcement-learning (RL)-based reward alignment and few-step distillation, yet these procedures ar