4 citations · 4 across the 1 of their papers we have counts for
1 paper
Yijie Lin, Jie Zhang, Zhenyu Huang +3
Existing video-language studies mainly focus on learning short video clips, leaving long-term temporal dependencies rarely explored due to over-high computational cost of modeling…