1 paper · 1 filter
Zheng Qi, Chao Shang, Evangelia Spiliopoulou +1
Vision language models (VLMs) often generate hallucination, i.e., content that cannot be substantiated by either textual or visual inputs. Prior work primarily attributes this to o…