1 paper · 1 filter
Ali Shirali, Arash Nasr-Esfahany, Abdullah Alomar +3
Alignment with human preferences is commonly framed using a universal reward function, even though human preferences are inherently heterogeneous. We formalize this heterogeneity b…