8 papers
REDAgentBench: Executable Red Teaming and Faithful Measurement of LLM Agent Systems
Zixing Chen, Xingyuan Liu, Jie Zhu +6
Large language model (LLM) agents combine language-based reasoning with external tools to perform complex tasks. Adversarial inputs can exploit interactions between the agent and i…
Dual-Loop Self-Evolution via Verifiable Emotion Feedback for Multi-Turn Empathetic Dialogue
Yi Wei, Shuo Jiang, Huaixia Dou +5
Large language models have demonstrated conversational capabilities, yet empathetic competence remains challenging. Empathetic support is inherently multi-turn and path-dependent:…
FinGuard: Detecting Financial Regulatory Non-Compliance in LLM Interactions
Huaixia Dou, Jie Zhu, Minghao Wu +5
As large language models (LLMs) are increasingly deployed in financial services, a single non-compliant interaction can expose institutions to regulatory penalties and direct consu…
ESC-Skills: Discovering and Self-Evolving Skills for Emotional Support Conversations
Jie Zhu, Huaixia Dou, Shuo Jiang +5
Existing emotional support conversation (ESC) systems mainly rely on end-to-end response generation or coarse strategy supervision, offering limited interpretability and little sup…
Modeling Multiple Support Strategies within a Single Turn for Emotional Support Conversations
Jie Zhu, Huaixia Dou, Junhui Li +5
Emotional Support Conversation (ESC) aims to assist individuals experiencing distress by generating empathetic and supportive dialogue. While prior work typically assumes that each…
Evaluating, Synthesizing, and Enhancing for Customer Support Conversation
Jie Zhu, Huaixia Dou, Junhui Li +4
Effective customer support requires not only accurate problem solving but also structured and empathetic communication aligned with professional standards. However, existing dialog…