6 citations · 6 across the 1 of their papers we have counts for
1 paper
Kirolos Ataallah, Xiaoqian Shen, Eslam Abdelrahman +4
This paper introduces MiniGPT4-Video, a multimodal Large Language Model (LLM) designed specifically for video understanding. The model is capable of processing both temporal visual…