1 paper
Yili Li, Jing Yu, Keke Gai +3
Current text-video retrieval methods mainly rely on cross-modal matching between queries and videos to calculate their similarity scores, which are then sorted to obtain retrieval…