1 paper
Xinying Lin, Xuyang Liu, Yiyu Wang +2
Video large language models (VideoLLMs) show strong capability in video understanding, yet long-context inference is still dominated by massive redundant visual tokens in the prefi…