16 citations · 16 across the 1 of their papers we have counts for
1 paper
Zijian Gao, Jingyu Liu, Weiqi Sun +3
Modern video-text retrieval frameworks basically consist of three parts: video encoder, text encoder and the similarity head. With the success on both visual and textual representa…