5 citations · 6 across the 2 of their papers we have counts for
2 papers
cs.CV2022★ 5 cited
Masked Contrastive Pre-Training for Efficient Video-Text Retrieval
Fangxun Shu, Biaolong Chen, Yue Liao +6
We present a simple yet effective end-to-end Video-language Pre-training (VidLP) framework, Masked Contrastive Video-language Pretraining (MAC), for video-text retrieval tasks. Our…
cs.CV2021★ 1 cited
Multi-Granularity Network with Modal Attention for Dense Affective Understanding
Baoming Yan, Lin Wang, Ke Gao +5
Video affective understanding, which aims to predict the evoked expressions by the video content, is desired for video creation and recommendation. In the recent EEV challenge, a d…