1 paper
Xiaochen Yang, Hao Fang, Jiawei Kong +3
Although large vision-language models (LVLMs) have demonstrated remarkable capabilities, they are prone to hallucinations in multi-image tasks. We attribute this issue to limitatio…