1 paper
Maximilian Augustin, Yannic Neuhaus, Matthias Hein
Vision-language models (VLMs) are prone to object hallucinations, where they erroneously indicate the presenceof certain objects in an image. Existing benchmarks quantify hallucina…