2 citations · 2 across the 19 of their papers we have counts for
1 paper · 1 filter
Sihang Jia, Shuliang Liu, Songbo Yang +1
Large vision-language models (LVLMs) frequently generate content unsupported by visual inputs. Preliminary experiments show that visual evidence is primarily incorporated into answ…