3 citations · 3 across the 1 of their papers we have counts for
1 paper
Guoshun Nan, Rui Qiao, Yao Xiao +4
Video grounding aims to localize a moment from an untrimmed video for a given textual query. Existing approaches focus more on the alignment of visual and language stimuli with var…