1 paper · 1 filter
Chengbing Wang, Wuqiang Zheng, Yang Zhang +5
Large Language Models (LLMs) are increasingly deployed in human-centric applications, yet they often fail to provide substantive emotional support. While Reinforcement Learning (RL…