30 citations · 30 across the 5 of their papers we have counts for
1 paper · 1 filter
Weijue Bu, Guan Yuan, Guixian Zhang
Large Vision-Language Models (VLMs) often exhibit text inertia, where attention drifts from visual evidence toward linguistic priors, resulting in object hallucinations. Existing d…