6 papers
How Much Do LLMs Know About Chinese Zero Pronouns?
Yifei Li, Guanyi Chen, Tingting He
Zero Pronouns (ZPs) are a pervasive linguistic phenomenon in pro-drop languages such as Chinese and have long posed a challenge for natural language processing systems. Although La…
When Seekers Are Hard to Help: Evaluating Emotional Support Dialogue Systems in Worst-Case Interactions
Jiajie Yang, Yangchun Li, Guanyi Chen +3
Emotional Support Dialogue Systems (ESDSes) are increasingly evaluated and trained with LLM-simulated seekers. However, such simulated seekers often behave as cooperative, average-…
On the Robustness of Knowledge Editing for Detoxification
Ming Dong, Shiyi Tang, Ziyan Peng +2
Knowledge-Editing-based (KE-based) detoxification has emerged as a promising approach for mitigating harmful behaviours in Large Language Models. Existing evaluations, however, lar…
How Do People Quantify Naturally: Evidence from Mandarin Picture Description
Yayun Zhang, Guanyi Chen, Fahime Same +2
Quantification is a fundamental component of everyday language use, yet little is known about how speakers decide whether and how to quantify in naturalistic production. We investi…
Do Large Language Models Judge Error Severity Like Humans?
Diege Sun, Guanyi Chen, Zhao Fan +2
Large Language Models (LLMs) are increasingly used as automated evaluators in natural language generation, yet it remains unclear whether they can accurately replicate human judgme…
Emotional Supporters often Use Multiple Strategies in a Single Turn
Xin Bai, Guanyi Chen, Tingting He +2
Emotional Support Conversations (ESC) are crucial for providing empathy, validation, and actionable guidance to individuals in distress. However, existing definitions of the ESC ta…