68 citations · 73 across the 2 of their papers we have counts for
2 papers
cs.CV2022★ 5 cited
Masked Contrastive Pre-Training for Efficient Video-Text Retrieval
Fangxun Shu, Biaolong Chen, Yue Liao +6
We present a simple yet effective end-to-end Video-language Pre-training (VidLP) framework, Masked Contrastive Video-language Pretraining (MAC), for video-text retrieval tasks. Our…
cs.CV2020★ 68 cited
Convolutional Hierarchical Attention Network for Query-Focused Video Summarization
Shuwen Xiao, Zhou Zhao, Zijian Zhang +2
Previous approaches for video summarization mainly concentrate on finding the most diverse and representative visual contents as video summary without considering the user's prefer…