1 paper
Kevin Clark, Paul Vicol, Kevin Swersky +1
We present Direct Reward Fine-Tuning (DRaFT), a simple and effective method for fine-tuning diffusion models to maximize differentiable reward functions, such as scores from human…