Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
Suppressing Prior-Comparison Hallucinations in Radiology Report Generation via Semantically Decoupled Latent Steering
Ao Li, Rui Liu, Mingjie Li +6
Automated radiology report generation using vision-language models (VLMs) is limited by the risk of prior-comparison hallucination, where the model generates historical findings un…
cs.CV2025
MedTrinity-25M: A Large-scale Multimodal Dataset with Multigranular Annotations for Medicine
Yunfei Xie, Ce Zhou, Lang Gao +8
This paper introduces MedTrinity-25M, a comprehensive, large-scale multimodal dataset for medicine, covering over 25 million images across 10 modalities with multigranular annotati…
cs.CV2024
Reducing Hallucinations in Vision-Language Models via Latent Space Steering
Sheng Liu, Haotian Ye, Lei Xing +1
Hallucination poses a challenge to the deployment of large vision-language models (LVLMs) in applications. Unlike in large language models (LLMs), hallucination in LVLMs often aris…