1 paper
Akhil Ramachandran, Ankit Arun, Ashish Shenoy +8
Video Large Language Models (Video LLMs) have shown remarkable progress in understanding and reasoning about visual content, particularly in tasks involving text recognition and te…