101 citations · 230 across the 18 of their papers we have counts for
13 papers · 1 filter
RCL: Recurrent Continuous Localization for Temporal Action Detection
Qiang Wang, Yanhao Zhang, Yun Zheng +1
Temporal representation is the cornerstone of modern action detection techniques. State-of-the-art methods mostly rely on a dense anchoring scheme, where anchors are sampled unifor…
Disentangled Representation Learning for Text-Video Retrieval
Qiang Wang, Yanhao Zhang, Yun Zheng +2
Cross-modality interaction is a critical component in Text-Video Retrieval (TVR), yet there has been little examination of how different influencing factors for computing interacti…
Learning Position and Target Consistency for Memory-based Video Object Segmentation
Li Hu, Peng Zhang, Bang Zhang +3
This paper studies the problem of semi-supervised video object segmentation(VOS). Multiple works have shown that memory-based approaches can be effective for video object segmentat…
Multiple Object Tracking with Correlation Learning
Qiang Wang, Yun Zheng, Pan Pan +1
Recent works have shown that convolutional networks have substantially improved the performance of multiple object tracking by simultaneously learning detection and appearance feat…
Few-Shot Incremental Learning with Continually Evolved Classifiers
Chi Zhang, Nan Song, Guosheng Lin +3
Few-shot class-incremental learning (FSCIL) aims to design machine learning algorithms that can continually learn new concepts from a few data points, without forgetting knowledge…
Self-supervised Video Representation Learning by Context and Motion Decoupling
Lianghua Huang, Yu Liu, Bin Wang +3
A key challenge in self-supervised video representation learning is how to effectively capture motion information besides context bias. While most existing works implicitly achieve…