3 papers
cs.CV2026
Role-Break in Attention Heads: Understanding and Detecting Hallucinations in VLMs
Mingyu Wang, Weilin Jin, Wenbo Li +5
Despite remarkable progress in vision-language generation, Vision-Language Models (VLMs) remain prone to hallucinations, producing content that is inconsistent with or unsupported…
cs.CV2026
HalluScope: Fine-grained Hallucination Diagnosis for Multimodal Large Language Models
Weilin Jin, Mingyu Wang, Wenbo Li +5
Although Multimodal Large Language Models have achieved strong performance across a wide range of vision-language tasks, they still suffer from hallucinations, where model outputs…
cs.CV2026
Geo3R: Mitigating Spatial Reasoning Hallucination in Multimodal Large Language Models
Mingyu Wang, Weilin Jin, Wenbo Li +3
Despite remarkable progress in visual understanding, Multimodal Large Language Models (MLLMs) remain prone to hallucinations when reasoning about spatial relationships, often produ…