2 papers
cs.CL2026
First, Do No Harm: AI Supervisor Scaffolds Novice Growth in Counselor Education
Chen Xu, Zhenyu Lyu, Tian Lan +12
The most dangerous mistakes a novice counselor makes are not the obvious ones: they are utterances that sound caring while quietly violating professional ethics and leaving vulnera…
cs.AI2025
Structural Reward Model: Enhancing Interpretability, Efficiency, and Scalability in Reward Modeling
Xiaoyu Liu, Di Liang, Chang Dai +9
Reward Models (RMs) are key components for evaluating and guiding language model outputs. However, traditional scalar RMs often struggle with incorporating contextual and backgroun…