RetinaFace: Single-stage Dense Face Localisation in the Wild
arXiv:1905.00641
Abstract
Though tremendous strides have been made in uncontrolled face detection, accurate and efficient face localisation in the wild remains an open challenge. This paper presents a robust single-stage face detector, named RetinaFace, which performs pixel-wise face localisation on various scales of faces by taking advantages of joint extra-supervised and self-supervised multi-task learning. Specifically, We make contributions in the following five aspects: (1) We manually annotate five facial landmarks on the WIDER FACE dataset and observe significant improvement in hard face detection with the assistance of this extra supervision signal. (2) We further add a self-supervised mesh decoder branch for predicting a pixel-wise 3D shape face information in parallel with the existing supervised branches. (3) On the WIDER FACE hard test set, RetinaFace outperforms the state of the art average precision (AP) by 1.1% (achieving AP equal to 91.4%). (4) On the IJB-C test set, RetinaFace enables state of the art methods (ArcFace) to improve their results in face verification (TAR=89.59% for FAR=1e-6). (5) By employing light-weight backbone networks, RetinaFace can run real-time on a single CPU core for a VGA-resolution image. Extra annotations and code have been made available at: https://github.com/deepinsight/insightface/tree/master/RetinaFace.
References in corpus (7)
- YOLO9000: Better, Faster, Stronger
- Face Attention Network: An Effective Face Detector for the Occluded Faces
- Face Detection through Scale-Friendly Deep Convolutional Networks
- Detecting Faces Using Region-based Fully Convolutional Networks
- Improved Selective Refinement Network for Face Detection
- WIDER Face and Pedestrian Challenge 2018: Methods and Results
- Robust and High Performance Face Detector
Cited by in corpus (42)
- TinaFace: Strong but Simple Baseline for Face Detection
- Deepfake Detection using Spatiotemporal Convolutional Networks
- Gendered Differences in Face Recognition Accuracy Explained by Hairstyles, Makeup, and Facial Morphology
- Going Deeper Into Face Detection: A Survey
- Large image datasets: A pyrrhic win for computer vision?
- CelebA-Spoof: Large-Scale Face Anti-Spoofing Dataset with Rich Annotations
- Achieving Better Kinship Recognition Through Better Baseline
- Multi-attentional Deepfake Detection
- CenterFace: Joint Face Detection and Alignment Using Face as Point
- ASFD: Automatic and Scalable Face Detector
- MaskFace: multi-task face and landmark detector
- CN-CELEB: a challenging Chinese speaker recognition dataset
- Light Can Hack Your Face! Black-box Backdoor Attack on Face Recognition Systems
- Deep Face Recognition Model Compression via Knowledge Transfer and Distillation
- Automated Anonymisation of Visual and Audio Data in Classroom Studies
- Face Detection with Feature Pyramids and Landmarks
- SMOT: Single-Shot Multi Object Tracking
- Emotion Recognition for In-the-wild Videos
- LDC-Net: A Unified Framework for Localization, Detection and Counting in Dense Crowds
- VarGFaceNet: An Efficient Variable Group Convolutional Neural Network for Lightweight Face Recognition
- T: Multi-Modal Continuous Valence-Arousal Estimation in the Wild
- Discriminability Distillation in Group Representation Learning
- Built Year Prediction from Buddha Face with Heterogeneous Labels
- AU-Guided Unsupervised Domain Adaptive Facial Expression Recognition
- An Effective Deep Network for Head Pose Estimation without Keypoints
- An Efficient Method for Face Quality Assessment on the Edge
- A 3D GAN for Improved Large-pose Facial Recognition
- Cognitive networks identify the content of English and Italian popular posts about COVID-19 vaccines: Anticipation, logistics, conspiracy and loss of trust
- Sub-pixel face landmarks using heatmaps and a bag of tricks
- More Information Supervised Probabilistic Deep Face Embedding Learning
- Advances and Challenges in Deep Lip Reading
- Disentangle Your Dense Object Detector
- Determination of the most representative descriptor among a set of feature vectors for the same object
- On Recognizing Occluded Faces in the Wild
- 3D High-Fidelity Mask Face Presentation Attack Detection Challenge
- When Face Recognition Meets Occlusion: A New Benchmark
- Audio-visual Speaker Recognition with a Cross-modal Discriminative Network
- TAL EmotioNet Challenge 2020 Rethinking the Model Chosen Problem in Multi-Task Learning
- Balanced Masked and Standard Face Recognition
- Face-NMS: A Core-set Selection Approach for Efficient Face Recognition
- Evaluation of Human and Machine Face Detection using a Novel Distinctive Human Appearance Dataset
- STC speaker recognition systems for the NIST SRE 2021