2 papers
cs.LG2025
APLOT: Robust Reward Modeling via Adaptive Preference Learning with Optimal Transport
Zhuo Li, Yuege Feng, Dandan Guo +3
The reward model (RM) plays a crucial role in aligning Large Language Models (LLMs) with human preferences through Reinforcement Learning, where the Bradley-Terry (BT) objective ha…
cs.CL2025
AgentMental: An Interactive Multi-Agent Framework for Explainable and Adaptive Mental Health Assessment
Jinpeng Hu, Ao Wang, Qianqian Xie +3
Mental health assessment is crucial for early intervention and effective treatment, yet traditional clinician-based approaches are limited by the shortage of qualified professional…