3 papers
cs.CL2026
Fin-PRM: A Domain-Specialized Process Reward Model for Financial Reasoning in Large Language Models
Jie Zhu, Yuanchen Zhou, Shuo Jiang +4
Process Reward Models (PRMs) supervise intermediate reasoning steps in large language models (LLMs), but existing PRMs are mainly trained on general-domain data and struggle with t…
hep-ph2026
The Pareto Frontier of Resilient Jet Tagging
Rikab Gambhir, Matt LeBlanc, Yuanchen Zhou
Classifying hadronic jets using their constituents' kinematic information is a critical task in modern high-energy collider physics. Often, classifiers are designed by targeting th…
cs.CL2026
CARE: Cognitive-reasoning Augmented Reinforcement for Emotional Support Conversation
Jie Zhu, Yuanchen Zhou, Shuo Jiang +5
Emotional Support Conversation (ESC) plays a vital role in alleviating psychological stress and providing emotional value through dialogue. While recent studies have largely focuse…