1 paper
Jeongwoo Shin, Dongsoo Shin, Yuchen Zhu +5
Reward fine-tuning has become a common approach for aligning pretrained diffusion and flow models with human preferences in text-to-image generation. Among reward-gradient-based me…