2 papers
cs.AI2025
Reflection-Bench: Evaluating Epistemic Agency in Large Language Models
Lingyu Li, Yixu Wang, Haiquan Zhao +4
With large language models (LLMs) increasingly deployed as cognitive engines for AI agents, the reliability and effectiveness critically hinge on their intrinsic epistemic agency,…
cs.CL2024
ESC-Eval: Evaluating Emotion Support Conversations in Large Language Models
Haiquan Zhao, Lingyu Li, Shisong Chen +11
Emotion Support Conversation (ESC) is a crucial application, which aims to reduce human stress, offer emotional guidance, and ultimately enhance human mental and physical well-bein…