167 citations · 167 across the 1 of their papers we have counts for
4 papers
Contrastive Learning of Image Representations with Cross-Video Cycle-Consistency
Haiping Wu, Xiaolong Wang
Recent works have advanced the performance of self-supervised representation learning by a large margin. The core among these methods is intra-image invariance learning. Two differ…
CvT: Introducing Convolutions to Vision Transformers
Haiping Wu, Bin Xiao, Noel Codella +4
We present in this paper a new architecture, named Convolutional vision Transformer (CvT), that improves Vision Transformer (ViT) in performance and efficiency by introducing convo…
Sequence Level Semantics Aggregation for Video Object Detection
Haiping Wu, Yuntao Chen, Naiyan Wang +1
Video objection detection (VID) has been a rising research direction in recent years. A central issue of VID is the appearance degradation of video frames caused by fast motion. Th…
Simple Baselines for Human Pose Estimation and Tracking
Bin Xiao, Haiping Wu, Yichen Wei
There has been significant progress on pose estimation and increasing interests on pose tracking in recent years. At the same time, the overall algorithm and system complexity incr…