1 paper
Jianfeng Dong, Lei Huang, Daizong Liu +5
Almost all previous text-to-video retrieval works ideally assume that videos are pre-trimmed with short durations containing solely text-related content. However, in practice, vide…