1 paper
Yansong Guo, Chaoyang Zhu, Jiayi Ji +2
Video Large Language Models (VideoLLMs) have demonstrated impressive capabilities in video understanding, yet the massive number of input video tokens incurs a significant computat…