3 citations · 5 across the 10 of their papers we have counts for
1 paper · 2 filters
Xiaofeng Zhang, Yihao Quan, Chen Shen +7
Large Vision Language Models (LVLMs) achieve great performance on visual-language reasoning tasks, however, the black-box nature of LVLMs hinders in-depth research on the reasoning…