1 paper
Fan Yuan, Chi Qin, Xiaogang Xu +1
Large Vision-Language Models (LVLMs) have shown remarkable performance on many visual-language tasks. However, these models still suffer from multimodal hallucination, which means…