Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
Meet Dynamic Individual Preferences: Resolving Conflicting Human Value with Paired Fine-Tuning
Shanyong Wang, Shuhang Lin, Yining Zhao +2
Recent advances in large language models (LLMs) have significantly improved the alignment of models with general human preferences. However, a major challenge remains in adapting L…
cs.CL2025
Sotopia-RL: Reward Design for Social Intelligence
Haofei Yu, Zhengyang Qi, Yining Zhao +6
Social intelligence has become a critical capability for large language models (LLMs), enabling them to engage effectively in real-world social tasks such as collaboration and nego…