1 paper · 1 filter
Kwangwook Seo, Dongha Lee
Recent approaches in personalized reward modeling have primarily focused on leveraging user interaction history to align model judgments with individual preferences. However, exist…