2 papers
cs.CL2026
BranPO: Scalable Contrastive Branch Sampling for Long-Horizon Agentic Reinforcement Learning
Yubao Zhao, Weiquan Huang, Sudong Wang +4
Agentic reinforcement learning enables large language models to perform multi-turn planning and tool use, but long-horizon training remains challenging under sparse trajectory-leve…
eess.SP2025
ECG-Chat: A Large ECG-Language Model for Cardiac Disease Diagnosis
Yubao Zhao, Jiaju Kang, Tian Zhang +2
The success of Multimodal Large Language Models (MLLMs) in the medical auxiliary field shows great potential, allowing patients to engage in conversations using physiological signa…