1 paper
Chenxi Li, Yichen Guo, Benfang Qian +5
Large Vision-Language Models (LVLMs) have achieved impressive performance in multimodal tasks, but they still suffer from hallucinations, i.e., generating content that is grammatic…