1 citations · 1 across the 4 of their papers we have counts for
1 paper · 1 filter
Chenglin Li, Qianglong Chen, Feng Han +6
Long-form video understanding remains a fundamental challenge for current Video Large Language Models. Most existing models rely on static reasoning over uniformly sampled frames,…