3 papers
cs.AI2026
Thinking with Deltas: Incentivizing Reinforcement Learning via Differential Visual Reasoning Policy
Shujian Gao, Yuan Wang, Jiangtao Yan +2
Reinforcement Learning with Verifiable Rewards (RLVR) has significantly advanced reasoning capabilities in Large Language Models. However, adapting RLVR to multimodal domains suffe…
cs.CE2025
Beyond N-grams: A Hierarchical Reward Learning Framework for Clinically-Aware Medical Report Generation
Yuan Wang, Shujian Gao, Jiaxiang Liu +6
Automatic medical report generation can greatly reduce the workload of doctors, but it is often unreliable for real-world deployment. Current methods can write formally fluent sent…
cs.CV2025
BARL: Bilateral Alignment in Representation and Label Spaces for Semi-Supervised Volumetric Medical Image Segmentation
Shujian Gao, Yuan Wang, Zekuan Yu
Semi-supervised medical image segmentation (SSMIS) seeks to match fully supervised performance while sharply reducing annotation cost. Mainstream SSMIS methods rely on \emph{label-…