Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
PERM: Psychology-grounded Empathetic Reward Modeling for Large Language Models
Chengbing Wang, Wuqiang Zheng, Yang Zhang +5
Large Language Models (LLMs) are increasingly deployed in human-centric applications, yet they often fail to provide substantive emotional support. While Reinforcement Learning (RL…
cs.CL2025
Think-While-Generating: On-the-Fly Reasoning for Personalized Long-Form Generation
Chengbing Wang, Yang Zhang, Wenjie Wang +4
Preference alignment has enabled large language models (LLMs) to better reflect human expectations, but current methods mostly optimize for population-level preferences, overlookin…