234 citations · 511 across the 12 of their papers we have counts for
12 papers · 1 filter
Dressing in the Wild by Watching Dance Videos
Xin Dong, Fuwei Zhao, Zhenyu Xie +6
While significant progress has been made in garment transfer, one of the most applicable directions of human-centric image generation, existing works overlook the in-the-wild image…
PP-YOLOv2: A Practical Object Detector
Xin Huang, Xinxin Wang, Wenyu Lv +10
Being effective and efficient is essential to an object detector for practical use. To meet these two concerns, we comprehensively evaluate a collection of existing refinements to…
RSPNet: Relative Speed Perception for Unsupervised Video Representation Learning
Peihao Chen, Deng Huang, Dongliang He +5
We study unsupervised video representation learning that seeks to learn both motion and appearance features from unlabeled video only, which can be reused for downstream tasks such…
PP-YOLO: An Effective and Efficient Implementation of Object Detector
Xiang Long, Kaipeng Deng, Guanzhong Wang +8
Object detection is one of the most important areas in computer vision, which plays a key role in various practical scenarios. Due to limitation of hardware, it is often necessary…
Graph-PCNN: Two Stage Human Pose Estimation with Graph Pose Refinement
Jian Wang, Xiang Long, Yuan Gao +2
Recently, most of the state-of-the-art human pose estimation methods are based on heatmap regression. The final coordinates of keypoints are obtained by decoding heatmap directly.…
Cross-Modality Attention with Semantic Graph Embedding for Multi-Label Classification
Renchun You, Zhiyao Guo, Lei Cui +3
Multi-label image and video classification are fundamental yet challenging tasks in computer vision. The main challenges lie in capturing spatial or temporal dependencies between l…