1 paper · 1 filter
Weijue Bu, Guan Yuan, Guixian Zhang
Large Vision-Language Models (VLMs) often exhibit text inertia, where attention drifts from visual evidence toward linguistic priors, resulting in object hallucinations. Existing d…