7 citations · 7 across the 3 of their papers we have counts for
1 paper · 1 filter
Yuchen Bao, Chao Wen, Haowei Wang +10
Reward post-training of diffusion generators inevitably concentrates probability mass on a few reward-favored modes, a mode collapse that erases within-prompt diversity. Existing m…