11 citations · 19 across the 6 of their papers we have counts for
1 paper · 1 filter
Satya Krishna Gorti, Noel Vouitsis, Junwei Ma +4
In text-video retrieval, the objective is to learn a cross-modal similarity function between a text and a video that ranks relevant text-video pairs higher than irrelevant pairs. H…