1 paper
Zhiwei Jia, Yuesong Nan, Huixi Zhao +1
Recent research has shown that fine-tuning diffusion models (DMs) with arbitrary rewards, including non-differentiable ones, is feasible with reinforcement learning (RL) techniques…