collaborators

10 papers

cs.CL2026

Web-CogReasoner: Towards Multimodal Knowledge-Induced Cognitive Reasoning for Web Agents

Yuhan Guo, Cong Guo, Aiwen Sun +12

Multimodal large-scale models have significantly advanced the development of web agents, enabling perception and interaction with digital environments akin to human cognition. In t…

cs.CL2026

Staying with the Uncertainty: Uncertainty-Scaffolding Strategies for Artificial Moral Advisors in LLM-to-LLM Simulated Conversations

Salvatore Greco, Hainiu Xu, Jacopo Domenicucci +2

LLMs are increasingly deployed as Artificial Moral Advisors (AMA) in a variety of contexts: what kind of conversational patterns should they display? In this paper, we study how AM…

cs.LG2026

PreAct-Bench: Benchmarking Predictive Monitoring in LLMs

Hainiu Xu, Italo Luis da Silva, Jiangnan Ye +7

Large language models (LLMs) are increasingly deployed as autonomous agents capable of executing multi-step action trajectories toward a given objective. While existing safety rese…

cs.HC2026

Why It Hurts: Identifying the Drivers of Negative Thoughts in Emotional Support Conversations

Hainiu Xu, Zhaoyue Sun, Hanqi Yan +3

Large Language Models (LLMs) are increasingly used for emotional support tasks, such as negative thought reframing. This task relies on modifying cognitive appraisals, the subjecti…

cs.AI2026

CAREBench: Evaluating LLMs' Emotion Understanding by Assessing Cognitive Appraisal Reasoning

Zhaoyue Sun, Hainiu Xu, Andero Uusberg +3

Emotion understanding is a core capability for LLMs to interact effectively with humans, yet existing evaluation paradigms rely on discrete emotion label prediction and fail to cap…

cs.CL2026

When Thinking Backfires: Mechanistic Insights Into Reasoning-Induced Misalignment

Hanqi Yan, Hainiu Xu, Siya Qi +2

With the growing accessibility and wide adoption of large language models, concerns about their safety and alignment with human values have become paramount. In this paper, we iden…