1 paper
Sreetama Sarkar, Yue Che, Alex Gavin +2
Despite their remarkable progress in multimodal understanding tasks, large vision language models (LVLMs) often suffer from "hallucinations", generating texts misaligned with the v…