1 citations · 1 across the 4 of their papers we have counts for
1 paper · 1 filter
Junpeng Ma, Qizhe Zhang, Ming Lu +4
Video Large Language Models (VLLMs) excel in video understanding, but their excessive visual tokens pose a significant computational challenge for real-world applications. Current…