1 paper · 1 filter
Amirabbas Afzali, Myeongho Jeon, Maria Brbic
Preference alignment is an essential step in adapting large language models (LLMs) to human values, but existing approaches typically depend on costly human annotations or large-sc…