activity
20242026
most citedSocialSim: Towards Socialized Simulation of Emotional Support Conversation

5 citations · 5 across the 6 of their papers we have counts for

collaborators

6 papers

cs.CL2026

PsychePass: Calibrating LLM Therapeutic Competence via Trajectory-Anchored Tournaments

Zhuang Chen, Dazhen Wan, Zhangkai Zheng +4

While large language models show promise in mental healthcare, evaluating their therapeutic competence remains challenging due to the unstructured and longitudinal nature of counse…

cs.CL2025

Unveiling the Landscape of Clinical Depression Assessment: From Behavioral Signatures to Psychiatric Reasoning

Zhuang Chen, Guanqun Bi, Wen Zhang +5

Depression is a widespread mental disorder that affects millions worldwide. While automated depression assessment shows promise, most studies rely on limited or non-clinically vali…

cs.CL20255 cited

SocialSim: Towards Socialized Simulation of Emotional Support Conversation

Zhuang Chen, Yaru Cao, Guanqun Bi +6

Emotional support conversation (ESC) helps reduce people's psychological stress and provide emotional value through interactive dialogues. Due to the high cost of crowdsourcing a l…

cs.CL2025

Ψ-Arena: Interactive Assessment and Optimization of LLM-based Psychological Counselors with Tripartite Feedback

Shijing Zhu, Zhuang Chen, Guanqun Bi +10

Large language models (LLMs) have shown promise in providing scalable mental health support, while evaluating their counseling capability remains crucial to ensure both efficacy an…

cs.CL2025

MAGI: Multi-Agent Guided Interview for Psychiatric Assessment

Guanqun Bi, Zhuang Chen, Zhoufu Liu +9

Automating structured clinical interviews could revolutionize mental healthcare accessibility, yet existing large language models (LLMs) approaches fail to align with psychiatric d…

cs.CL2024

CharacterBench: Benchmarking Character Customization of Large Language Models

Jinfeng Zhou, Yongkang Huang, Bosi Wen +13

Character-based dialogue (aka role-playing) enables users to freely customize characters for interaction, which often relies on LLMs, raising the need to evaluate LLMs' character c…