17 citations · 22 across the 4 of their papers we have counts for
1 paper · 2 filters
Zhu Zhang, Zhou Zhao, Yang Zhao +3
In this paper, we consider a novel task, Spatio-Temporal Video Grounding for Multi-Form Sentences (STVG). Given an untrimmed video and a declarative/interrogative sentence depictin…