4 citations · 6 across the 2 of their papers we have counts for
3 papers
cs.CV2022★ 4 cited
SMAUG: Sparse Masked Autoencoder for Efficient Video-Language Pre-training
Yuanze Lin, Chen Wei, Huiyu Wang +2
Video-language pre-training is crucial for learning powerful multi-modal representation. However, it typically requires a massive amount of computation. In this paper, we develop S…
cs.CV2021★ 2 cited
Self-Supervised Video Representation Learning with Meta-Contrastive Network
Yuanze Lin, Xun Guo, Yan Lu
Self-supervised learning has been successfully applied to pre-train video representations, which aims at efficient adaptation from pre-training domain to downstream tasks. Existing…
cs.CV2019
An Action Recognition network for specific target based on rMC and RPN
Mingjie Li, Youqian Feng, Zhonghai Yin +4
The traditional methods of action recognition are not specific for the operator, thus results are easy to be disturbed when other actions are operated in videos. The network based…