3 papers
cs.LG2026
Understanding the Role of Hallucination in Reinforcement Post-Training of Multimodal Reasoning Models
Gengwei Zhang, Jie Peng, Zhen Tan +6
The recent success of reinforcement learning (RL) in large reasoning models has inspired the growing adoption of RL for post-training Multimodal Large Language Models (MLLMs) to en…
cs.CV2025
PPBoost: Progressive Prompt Boosting for Text-Driven Medical Image Segmentation
Xuchen Li, Hengrui Gu, Mohan Zhang +6
Text-prompted foundation models for medical image segmentation offer an intuitive way to delineate anatomical structures from natural language queries, but their predictions often…
cs.CV2025
Fairness in Multi-modal Medical Diagnosis with Demonstration Selection
Dawei Li, Zijian Gu, Peng Wang +6
Multimodal large language models (MLLMs) have shown strong potential for medical image reasoning, yet fairness across demographic groups remains a major concern. Existing debiasing…