3 citations · 10 across the 8 of their papers we have counts for
1 paper · 1 filter
Haotian Ye, Kaiwen Zheng, Jiashu Xu +15
Aligning generative diffusion models with human preferences via reinforcement learning (RL) is critical yet challenging. Most existing algorithms are often vulnerable to reward hac…