3 papers
cs.CV2025
Anatomical Region-Guided Contrastive Decoding: A Plug-and-Play Strategy for Mitigating Hallucinations in Medical VLMs
Xiao Liang, Chenxi Liu, Zhi Ma +4
Medical Vision-Language Models (MedVLMs) show immense promise in clinical applicability. However, their reliability is hindered by hallucinations, where models often fail to derive…
cs.AI2025
VideoAgent: Personalized Synthesis of Scientific Videos
Xiao Liang, Bangxin Li, Zixuan Chen +5
The technical complexity of research papers often limits their reach, necessitating more accessible formats like scientific videos to disseminate key insights through engaging narr…
cs.CV2025
CheXPO: Preference Optimization for Chest X-ray VLMs with Counterfactual Rationale
Xiao Liang, Jiawei Hu, Di Wang +5
Vision-language models (VLMs) are prone to hallucinations that critically compromise reliability in medical applications. While preference optimization can mitigate these hallucina…