1 citations · 1 across the 2 of their papers we have counts for
1 paper · 1 filter
Louis Castricato, Nathan Lile, Rafael Rafailov +2
The rapid advancement of language models (LMs) necessitates robust alignment with diverse user values. However, current preference optimization approaches often fail to capture the…