Preprint
Aug 2026
Scaling Reinforcement Learning for Diffusion Models via Velocity Matching
This work proposes reward-based velocity matching (RVM), a simple trajectory-free update that acts directly on the velocity field and provides a general framework that recovers recent fine-tuning methods, including RAM and DiffusionNFT, as special cases.
Jaemoo Choi, Wei Guo, Yuchen Zhu et al.
· 0 citations