1 paper
Zongmeng Zhang, Wengang Zhou, Jie Zhao +1
Despite the impressive capabilities of multimodal large language models (MLLMs) in vision-language tasks, they are prone to hallucinations in real-world scenarios. This paper inves…