11 citations · 11 across the 1 of their papers we have counts for
1 paper
Satya Krishna Gorti, Noel Vouitsis, Junwei Ma +4
In text-video retrieval, the objective is to learn a cross-modal similarity function between a text and a video that ranks relevant text-video pairs higher than irrelevant pairs. H…