1 paper
Tingyu Qu, Mingxiao Li, Tinne Tuytelaars +1
Recent advances in multimodal Large Language Models (LLMs) have shown great success in understanding multi-modal contents. For video understanding tasks, training-based video LLMs…