1 citations · 1 across the 4 of their papers we have counts for
1 paper · 1 filter
Siyou Li, Huanan Wu, Juexi Shao +10
Despite the recent advances in the video understanding ability of multimodal large language models (MLLMs), long video understanding remains a challenge. One of the main issues is…