Single Person Pose Estimation: A Survey
arXiv:2109.10056
Abstract
Human pose estimation in unconstrained images and videos is a fundamental computer vision task. To illustrate the evolutionary path in technique, in this survey we summarize representative human pose methods in a structured taxonomy, with a particular focus on deep learning models and single-person image setting. Specifically, we examine and survey all the components of a typical human pose estimation pipeline, including data augmentation, model architecture and backbone, supervision representation, post-processing, standard datasets, evaluation metrics. To envisage the future directions, we finally discuss the key unsolved problems and potential trends for human pose estimation.
16 pages, 3 figures
References in corpus (9)
- Very Deep Convolutional Networks for Large-Scale Image Recognition
- Neural Architecture Search with Reinforcement Learning
- Joint Training of a Convolutional Network and a Graphical Model for Human Pose Estimation
- Monocular Human Pose Estimation: A Survey of Deep Learning-based Methods
- Articulated Pose Estimation by a Graphical Model with Image Dependent Pairwise Relations
- Human Pose Estimation with Spatial Contextual Information
- CRF-CNN: Modeling Structured Information in Human Pose Estimation
- Anti-Confusing: Region-Aware Network for Human Pose Estimation
- Train Your Data Processor: Distribution-Aware and Error-Compensation Coordinate Decoding for Human Pose Estimation