1 paper
Xueqing Yu, Bohan Li, Yan Li +1
Recent Vision-Language Models (VLMs) have made remarkable progress in multimodal understanding tasks, yet their evaluation on long video understanding remains unreliable. Due to li…