1 citations · 1 across the 3 of their papers we have counts for
1 paper · 1 filter
Yuanshen Guan, Zipeng Feng, Chengru Song +2
Aligning diffusion models with human preferences usually relies on a sparse terminal reward evaluated on the final generated samples, which creates a severe temporal credit-assignm…