7 citations · 7 across the 2 of their papers we have counts for
2 papers
cs.CV2021★ 7 cited
Video-Text Pre-training with Learned Regions
Rui Yan, Mike Zheng Shou, Yixiao Ge +4
Video-Text pre-training aims at learning transferable representations from large-scale video-text pairs via aligning the semantics between visual and textual information. State-of-…
cs.CV2021
Object-aware Video-language Pre-training for Retrieval
Alex Jinpeng Wang, Yixiao Ge, Guanyu Cai +5
Recently, by introducing large-scale dataset and strong transformer network, video-language pre-training has shown great success especially for retrieval. Yet, existing video-langu…