1 paper
Xiangtian Zheng, Zishuo Wang, Yuxin Peng
With the rapid development of Large Language Models (LLMs), Video Multi-Modal Large Language Models (Video MLLMs) have achieved remarkable performance in video-language tasks such…