3 papers
cs.CL2025
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models
Wenhui Zhu, Xuanzhao Dong, Xin Li +6
Recently, reinforcement learning (RL)-based tuning has shifted the trajectory of Multimodal Large Language Models (MLLMs), particularly following the introduction of Group Relative…
cs.CV2025
Schrödinger Diffusion Driven Signal Recovery in 3T BOLD fMRI Using Unmatched 7T Observations
Yujian Xiong, Xuanzhao Dong, Sebastian Waz +4
Ultra-high-field (7 Tesla) BOLD fMRI offers exceptional detail in both spatial and temporal domains, along with robust signal-to-noise characteristics, making it a powerful modalit…
cs.CV2024
Many-MobileNet: Multi-Model Augmentation for Robust Retinal Disease Classification
Hao Wang, Wenhui Zhu, Xuanzhao Dong +9
In this work, we propose Many-MobileNet, an efficient model fusion strategy for retinal disease classification using lightweight CNN architecture. Our method addresses key challeng…