17 citations · 18 across the 2 of their papers we have counts for
3 papers
cs.CV2022★ 1 cited
DR.VIC: Decomposition and Reasoning for Video Individual Counting
Tao Han, Lei Bai, Junyu Gao +2
Pedestrian counting is a fundamental tool for understanding pedestrian patterns and crowd flow analysis. Existing works (e.g., image-level pedestrian counting, crossline crowd coun…
cs.CV2020
Where Does It Exist: Spatio-Temporal Video Grounding for Multi-Form Sentences
Zhu Zhang, Zhou Zhao, Yang Zhao +3
In this paper, we consider a novel task, Spatio-Temporal Video Grounding for Multi-Form Sentences (STVG). Given an untrimmed video and a declarative/interrogative sentence depictin…
cs.CV2019★ 17 cited
Weakly-Supervised Video Moment Retrieval via Semantic Completion Network
Zhijie Lin, Zhou Zhao, Zhu Zhang +2
Video moment retrieval is to search the moment that is most relevant to the given natural language query. Existing methods are mostly trained in a fully-supervised setting, which r…