most citedFin-PRM: A Domain-Specialized Process Reward Model for Financial Reasoning in Large Language Models

2 citations · 2 across the 6 of their papers we have counts for

collaborators
Showing cs.CLShow all

5 papers · 1 filter

cs.CL2026

Dual-Loop Self-Evolution via Verifiable Emotion Feedback for Multi-Turn Empathetic Dialogue

Yi Wei, Shuo Jiang, Huaixia Dou +5

Large language models have demonstrated conversational capabilities, yet empathetic competence remains challenging. Empathetic support is inherently multi-turn and path-dependent:…

cs.CL2026

FinGuard: Detecting Financial Regulatory Non-Compliance in LLM Interactions

Huaixia Dou, Jie Zhu, Minghao Wu +5

As large language models (LLMs) are increasingly deployed in financial services, a single non-compliant interaction can expose institutions to regulatory penalties and direct consu…

cs.CL2026

ESC-Skills: Discovering and Self-Evolving Skills for Emotional Support Conversations

Jie Zhu, Huaixia Dou, Shuo Jiang +5

Existing emotional support conversation (ESC) systems mainly rely on end-to-end response generation or coarse strategy supervision, offering limited interpretability and little sup…

cs.CL20262 cited

Fin-PRM: A Domain-Specialized Process Reward Model for Financial Reasoning in Large Language Models

Jie Zhu, Yuanchen Zhou, Shuo Jiang +4

Process Reward Models (PRMs) supervise intermediate reasoning steps in large language models (LLMs), but existing PRMs are mainly trained on general-domain data and struggle with t…

cs.CL2026

CARE: Cognitive-reasoning Augmented Reinforcement for Emotional Support Conversation

Jie Zhu, Yuanchen Zhou, Shuo Jiang +5

Emotional Support Conversation (ESC) plays a vital role in alleviating psychological stress and providing emotional value through dialogue. While recent studies have largely focuse…