1 paper
Jie Ren, Yuhang Zhang, Dongrui Liu +2
Direct preference optimization (DPO) has shown success in aligning diffusion models with human preference. Previous approaches typically assume a consistent preference label betwee…