131 citations · 131 across the 1 of their papers we have counts for
1 paper
Han Fang, Pengfei Xiong, Luhui Xu +1
We present CLIP2Video network to transfer the image-language pre-training model to video-text retrieval in an end-to-end manner. Leading approaches in the domain of video-and-langu…