1 paper
Yijun Song, Jingwen Wang, Lin Ma +2
The task of temporally grounding textual queries in videos is to localize one video segment that semantically corresponds to the given query. Most of the existing approaches rely o…