Showing cs.CLShow all
3 papers · 1 filter
cs.CL2025
RAIDEN-R1: Improving Role-awareness of LLMs via GRPO with Verifiable Reward
Zongsheng Wang, Kaili Sun, Bowen Wu +3
Role-playing conversational agents (RPCAs) face persistent challenges in maintaining role consistency. To address this, we propose RAIDEN-R1, a novel reinforcement learning framewo…
cs.CL2025
Improving Generalization in Intent Detection: GRPO with Reward-Based Curriculum Sampling
Zihao Feng, Xiaoxue Wang, Ziwei Bai +4
Intent detection, a critical component in task-oriented dialogue (TOD) systems, faces significant challenges in adapting to the rapid influx of integrable tools with complex interr…
cs.CL2025
Interpersonal Memory Matters: A New Task for Proactive Dialogue Utilizing Conversational History
Bowen Wu, Wenqing Wang, Haoran Li +3
Proactive dialogue systems aim to empower chatbots with the capability of leading conversations towards specific targets, thereby enhancing user engagement and service autonomy. Ex…