14 citations · 70 across the 73 of their papers we have counts for
1 paper · 2 filters
Yixin Liu, Argyris Oikonomou, Weiqiang Zheng +2
Many alignment methods, including reinforcement learning from human feedback (RLHF), rely on the Bradley-Terry reward assumption, which is not always sufficient to capture the full…