1 paper
Binwei Yao, Zefan Cai, Yun-Shiuan Chuang +4
Preferences within a group of people are not uniform but follow a distribution. While existing alignment methods like Direct Preference Optimization (DPO) attempt to steer models t…