1 citations · 1 across the 1 of their papers we have counts for
1 paper
Yueqian Wang, Xiaojun Meng, Jianxin Liang +3
Video-text Large Language Models (video-text LLMs) have shown remarkable performance in answering questions and holding conversations on simple videos. However, they perform almost…