7 citations · 21 across the 17 of their papers we have counts for
1 paper · 1 filter
Shuo Liu, Weize Quan, Ming Zhou +5
Videos contain multi-modal content, and exploring multi-level cross-modal interactions with natural language queries can provide great prominence to text-video retrieval task (TVR)…