04:00
2026-08-26
arxiv.org
machine-learning
Scaling Reinforcement Learning for Diffusion Models via Velocity Matching
Researchers propose reward-based velocity matching (RVM), a trajectory-free update for fine-tuning diffusion models that acts directly on the velocity field, eliminating the need for likelihood-based โฆ