1 paper
Leon Mayer, Lucas Luttner, Patrick Godau +42
Recent advances in Vision-Language Models (VLMs) have led to rapid progress in video understanding across a wide range of benchmark tasks. However, existing evaluations largely foc…