178 citations · 870 across the 28 of their papers we have counts for
1 paper · 2 filters
Liang Chen, Haozhe Zhao, Tianyu Liu +4
In this study, we identify the inefficient attention phenomena in Large Vision-Language Models (LVLMs), notably within prominent models like LLaVA-1.5, QwenVL-Chat and Video-LLaVA.…