19 citations · 66 across the 15 of their papers we have counts for
4 papers · 1 filter
Hierarchical Local-Global Transformer for Temporal Sentence Grounding
Xiang Fang, Daizong Liu, Pan Zhou +2
This paper studies the multimedia problem of temporal sentence grounding (TSG), which aims to accurately determine the specific video segment in an untrimmed video according to a g…
Backdoor Attacks on Crowd Counting
Yuhua Sun, Tailai Zhang, Xingjun Ma +6
Crowd counting is a regression task that estimates the number of people in a scene image, which plays a vital role in a range of safety-critical applications, such as video surveil…
Unsupervised Temporal Video Grounding with Deep Semantic Clustering
Daizong Liu, Xiaoye Qu, Yinzhen Wang +5
Temporal video grounding (TVG) aims to localize a target segment in a video according to a given sentence query. Though respectable works have made decent achievements in this task…
Memory-Guided Semantic Learning Network for Temporal Sentence Grounding
Daizong Liu, Xiaoye Qu, Xing Di +3
Temporal sentence grounding (TSG) is crucial and fundamental for video understanding. Although the existing methods train well-designed deep networks with a large amount of data, w…