139 citations · 385 across the 8 of their papers we have counts for
10 papers · 1 filter
RCL: Recurrent Continuous Localization for Temporal Action Detection
Qiang Wang, Yanhao Zhang, Yun Zheng +1
Temporal representation is the cornerstone of modern action detection techniques. State-of-the-art methods mostly rely on a dense anchoring scheme, where anchors are sampled unifor…
Disentangled Representation Learning for Text-Video Retrieval
Qiang Wang, Yanhao Zhang, Yun Zheng +2
Cross-modality interaction is a critical component in Text-Video Retrieval (TVR), yet there has been little examination of how different influencing factors for computing interacti…
Multiple Object Tracking with Correlation Learning
Qiang Wang, Yun Zheng, Pan Pan +1
Recent works have shown that convolutional networks have substantially improved the performance of multiple object tracking by simultaneously learning detection and appearance feat…
Fashion Focus: Multi-modal Retrieval System for Video Commodity Localization in E-commerce
Yanhao Zhang, Qiang Wang, Pan Pan +4
Nowadays, live-stream and short video shopping in E-commerce have grown exponentially. However, the sellers are required to manually match images of the selling products to the tim…
Anti-UAV: A Large Multi-Modal Benchmark for UAV Tracking
Nan Jiang, Kuiran Wang, Xiaoke Peng +7
Unmanned Aerial Vehicle (UAV) offers lots of applications in both commerce and recreation. With this, monitoring the operation status of UAVs is crucially important. In this work,…
Anchor Diffusion for Unsupervised Video Object Segmentation
Zhao Yang, Qiang Wang, Luca Bertinetto +3
Unsupervised video object segmentation has often been tackled by methods based on recurrent neural networks and optical flow. Despite their complexity, these kinds of approaches te…