31 citations · 33 across the 5 of their papers we have counts for
1 paper · 1 filter
Xiaoye Qu, Qiyuan Chen, Wei Wei +2
Despite the remarkable ability of large vision-language models (LVLMs) in image comprehension, these models frequently generate plausible yet factually incorrect responses, a pheno…