1 paper
Shimin Chen, Yitian Yuan, Shaoxiang Chen +2
Amidst the advancements in image-based Large Vision-Language Models (image-LVLM), the transition to video-based models (video-LVLM) is hindered by the limited availability of quali…