1 paper
Xinxin Liu, Shiwei Gan, Xiao Liu +3
Video Large Language Models (Video-LLMs) achieve strong performance in video understanding, but their excessive visual tokens bring substantial computational overhead. Existing tra…