1 paper
Haichao Zhang, Yao Lu, Lichen Wang +4
Video Large Language Models (VLLMs) unlock world-knowledge-aware video understanding through pretraining on internet-scale data and have already shown promise on tasks such as movi…