WebFace260M: A Benchmark Unveiling the Power of Million-Scale Deep Face Recognition
Zheng Zhu, Guan Huang, Jiankang Deng, Yun Ye, Junjie Huang, Xinze Chen, Jiagang Zhu, Tian Yang, Jiwen Lu, Dalong Du, Jie Zhou
Abstract
In this paper, we contribute a new million-scale face benchmark containing noisy 4M identities/260M faces (WebFace260M) and cleaned 2M identities/42M faces (WebFace42M) training data, as well as an elaborately designed time-constrained evaluation protocol. Firstly, we collect 4M name list and download 260M faces from the Internet. Then, a Cleaning Automatically utilizing Self-Training (CAST) pipeline is devised to purify the tremendous WebFace260M, which is efficient and scalable. To the best of our knowledge, the cleaned WebFace42M is the largest public face recognition training set and we expect to close the data gap between academia and industry. Referring to practical scenarios, Face Recognition Under Inference Time conStraint (FRUITS) protocol and a test set are constructed to comprehensively evaluate face matchers.
Equipped with this benchmark, we delve into millionscale face recognition problems. A distributed framework is developed to train face recognition models efficiently without tampering with the performance. Empowered by Web-Face42M, we reduce relative 40% failure rate on the challenging IJB-C set, and ranks the 3rd among 430 entries on NIST-FRVT. Even 10% data (WebFace4M) shows superior performance compared with public training set. Furthermore, comprehensive baselines are established on our rich-attribute test set under FRUITS-100ms/500ms/1000ms protocol, including MobileNet, EfficientNet, AttentionNet, ResNet, SENet, ResNeXt and RegNet families. Benchmark website is https://www.face-benchmark.org.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 04ed06ba-929c-4006-b174-9e81256a8faaCited by top-tier papers58
- AdaFace: Quality Adaptive Margin for Face RecognitionMinchul Kim, Anil K. Jain, Xiaoming LiuCVPR 2022 · 509 citations
- Face2Exp: Combating Data Biases for Facial Expression RecognitionDan Zeng, Zhiyuan Lin, Xiao Yan, Yuting Liu et al.CVPR 2022 · 125 citations
- Gait Recognition in the Wild: A BenchmarkICCV 2021 · 102 citations
- BlendFace: Re-designing Identity Encoders for Face-SwappingKaede Shiohara, Xingchao Yang, Takafumi TaketomiICCV 2023 · 83 citations
- Killing Two Birds with One Stone: Efficient and Robust Training of Face Recognition CNNs by Partial FCXiang An, Jiankang Deng, Jia Guo, Ziyong Feng et al.CVPR 2022 · 77 citations
Builds on12
- Racial Faces in the Wild: Reducing Racial Bias by Information Maximization Adaptation NetworkMei Wang, Weihong Deng, Jiani Hu, Xunqiang Tao et al.ICCV 2019 · 379 citations
- Probabilistic Face EmbeddingsYichun Shi, Anil K. JainICCV 2019 · 362 citations
- RetinaFace: Single-Shot Multi-Level Face Localisation in the WildJiankang Deng, Jia Guo, Evangelos Ververas, Irene Kotsia et al.CVPR 2020
- Learning to Cluster Faces via Confidence and Connectivity EstimationLei Yang, Dapeng Chen, Xiaohang Zhan, Rui Zhao et al.CVPR 2020
- Mitigating Face Recognition Bias via Group Adaptive ClassifierSixue Gong, Xiaoming Liu, Anil K. JainCVPR 2021
Related papers
- Global-Local GCN: Large-Scale Label Noise Cleansing for Face RecognitionYaobin Zhang, Weihong Deng, Mei Wang, Jiani Hu et al.CVPR 2020
- How to Boost Face Recognition with StyleGAN?Artem Sevastopolsky, Yury Malkov, Nikita Durasov, Luisa Verdoliva et al.ICCV 2023 · 17 citations
- DeeperForensics-1.0: A Large-Scale Dataset for Real-World Face Forgery DetectionLiming Jiang, Ren Li, Wayne Wu, Chen Qian et al.CVPR 2020
- Stylized-Face: A Million-Level Stylized Face Dataset for Face RecognitionZhengyuan Peng, Jianqing Xu, Yuge Huang, Jinkun Hao et al.ICCV 2025 · 1 citation
- Learning Meta Face Recognition in Unseen DomainsJianzhu Guo, Xiangyu Zhu, Chenxu Zhao, Dong Cao et al.CVPR 2020
