1 paper · 1 filter
Liang Chen, Haozhe Zhao, Tianyu Liu +4
In this study, we identify the inefficient attention phenomena in Large Vision-Language Models (LVLMs), notably within prominent models like LLaVA-1.5, QwenVL-Chat and Video-LLaVA.…