7 citations · 7 across the 1 of their papers we have counts for
1 paper
Rui Yan, Mike Zheng Shou, Yixiao Ge +4
Video-Text pre-training aims at learning transferable representations from large-scale video-text pairs via aligning the semantics between visual and textual information. State-of-…