1 paper
Ali Shirali, Arash Nasr-Esfahany, Abdullah Alomar +3
Alignment with human preferences is commonly framed using a universal reward function, even though human preferences are inherently heterogeneous. We formalize this heterogeneity b…