1 paper · 1 filter
Woojin Kim, Sieun Hyeon, Jusang Oh +1
Aligning Large Language Models (LLMs) with the diverse spectrum of human values remains a central challenge: preference-based methods often fail to capture deeper motivational prin…