1 paper · 1 filter
Mohamad Chehade, Soumya Suvra Ghosal, Souradip Chakraborty +4
Aligning large language models with humans is challenging due to the inherently multifaceted nature of preference feedback. While existing approaches typically frame this as a mult…