1 paper · 1 filter
Yujie Zhou, Pengyang Ling, Jiazi Bu +5
The incorporation of online reinforcement learning (RL) into diffusion and flow-based generative models has recently gained attention as a powerful paradigm for aligning model beha…