MS-Celeb-1M: A Dataset and Benchmark for Large-Scale Face Recognition
arXiv:1607.08221
Abstract
In this paper, we design a benchmark task and provide the associated datasets for recognizing face images and link them to corresponding entity keys in a knowledge base. More specifically, we propose a benchmark task to recognize one million celebrities from their face images, by using all the possibly collected face images of this individual on the web as training data. The rich information provided by the knowledge base helps to conduct disambiguation and improve the recognition accuracy, and contributes to various real-world applications, such as image captioning and news video analysis. Associated with this task, we design and provide concrete measurement set, evaluation protocol, as well as training data. We also present in details our experiment setup and report promising baseline results. Our benchmark task could lead to one of the largest classification problems in computer vision. To the best of our knowledge, our training dataset, which contains 10M images in version 1, is the largest publicly available one in the world.
References in corpus (2)
Cited by in corpus (25)
- Deep Feature Augmentation for Occluded Image Classification
- von Mises-Fisher Mixture Model-based Deep learning: Application to Face Verification
- Range Loss for Deep Face Recognition with Long-tail
- Inducing Predictive Uncertainty Estimation for Face Recognition
- Mis-classified Vector Guided Softmax Loss for Face Recognition
- Adversarial Discriminative Heterogeneous Face Recognition
- Anti-Makeup: Learning A Bi-Level Adversarial Network for Makeup-Invariant Face Verification
- Masked Face Recognition Challenge: The InsightFace Track Report
- Video Face Recognition: Component-wise Feature Aggregation Network (C-FAN)
- DeepDeblur: Fast one-step blurry face images restoration
- Accelerated Training for Massive Classification via Dynamic Class Selection
- A Scalable Approach for Facial Action Unit Classifier Training UsingNoisy Data for Pre-Training
- UV-GAN: Adversarial Facial UV Map Completion for Pose-invariant Face Recognition
- SuperFront: From Low-resolution to High-resolution Frontal Face Synthesis
- Merge or Not? Learning to Group Faces via Imitation Learning
- Efficient Realistic Data Generation Framework leveraging Deep Learning-based Human Digitization
- Correlation Congruence for Knowledge Distillation
- Imponderous Net for Facial Expression Recognition in the Wild
- Large Scale Incremental Learning
- DotFAN: A Domain-transferred Face Augmentation Network for Pose and Illumination Invariant Face Recognition
- Learning to Cluster Faces on an Affinity Graph
- Visual Data Augmentation through Learning
- Automatically Building Face Datasets of New Domains from Weakly Labeled Data with Pretrained Models
- Large-scale Multi-modal Person Identification in Real Unconstrained Environments
- MixFaceNets: Extremely Efficient Face Recognition Networks