Killing Two Birds with One Stone: Efficient and Robust Training of Face Recognition CNNs by Partial FC
Xiang An, Jiankang Deng, Jia Guo, Ziyong Feng, Xuhan Zhu, Jing Yang, Tongliang Liu
Abstract
Learning discriminative deep feature embeddings by using million-scale in-the-wild datasets and margin-based softmax loss is the current state-of-the-art approach for face recognition. However, the memory and computing cost of the Fully Connected (FC) layer linearly scales up to the number of identities in the training set. Besides, the largescale training data inevitably suffers from inter-class conflict and long-tailed distribution. In this paper, we propose a sparsely updating variant of the FC layer, named Partial FC (PFC). In each iteration, positive class centers and a random subset of negative class centers are selected to compute the margin-based softmax loss. All class centers are still maintained throughout the whole training process, but only a subset is selected and updated in each iteration. Therefore, the computing requirement, the probability of inter-class conflict, and the frequency of passive update on tail class centers, are dramatically reduced. Extensive experiments across different training data and backbones (e.g. CNN and ViT) confirm the effectiveness, robustness and efficiency of the proposed PFC. The source code is available at https://github.com/deepinsight/ insightface/tree/master/recognition.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a477b9e9-1f49-4f9d-9fec-cbdf6d521469Cited by top-tier papers18
- UniFace: Unified Cross-Entropy Loss for Deep Face RecognitionJiancan Zhou, Xi Jia, Qiufu Li, Linlin Shen et al.ICCV 2023 · 38 citations
- TransFace: Calibrating Transformer Training for Face Recognition from a Data-Centric PerspectiveJun Dan, Yang Liu, Haoyu Xie, Jiankang Deng et al.ICCV 2023 · 36 citations
- Standing on the Shoulders of Giants: Reprogramming Visual-Language Model for General Deepfake DetectionKaiqing Lin, Yuzhen Lin, Weixiang Li, Taiping Yao et al.AAAI 2025 · 32 citations
- TopoFR: A Closer Look at Topology Alignment on Face RecognitionJun Dan, Yang Liu, Jiankang Deng, Haoyu Xie et al.NeurIPS 2024 · 27 citations
- CLIP-CID: Efficient CLIP Distillation via Cluster-Instance DiscriminationKaicheng Yang, Tiancheng Gu, Xiang An, Haiqiang Jiang et al.AAAI 2025 · 26 citations
Builds on18
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Mis-Classified Vector Guided Softmax Loss for Face RecognitionXiaobo Wang, Shifeng Zhang, Shuo Wang, Tianyu Fu et al.AAAI 2020 · 188 citations
- Co-Mining: Deep Face Recognition With Noisy LabelsXiaobo Wang, Shuo Wang, Hailin Shi, Jun Wang et al.ICCV 2019 · 114 citations
- Attentional Feature-Pair Relation Networks for Accurate Face RecognitionBong-Nam Kang, Yonghyun Kim, Bongjin Jun, Daijin KimICCV 2019 · 38 citations
- Softmax Dissection: Towards Understanding Intra- and Inter-Class Objective for Embedding LearningLanqing He, Zhongdao Wang, Yali Li, Shengjin WangAAAI 2020 · 34 citations
Related papers
- Dynamic Class Queue for Large Scale Face Recognition in the WildBi Li, Teng Xi, Gang Zhang, Haocheng Feng et al.CVPR 2021
- An Efficient Training Approach for Very Large Scale Face RecognitionKai Wang, Shuo Wang, Panpan Zhang, Zhipeng Zhou et al.CVPR 2022 · 29 citations
- Virtual Fully-Connected Layer: Training a Large-Scale Face Recognition Dataset With Limited Computational ResourcesPengyu Li, Biao Wang, Lei ZhangCVPR 2021
- Variational Prototype Learning for Deep Face RecognitionJiankang Deng, Jia Guo, Jing Yang, Alexandros Lattas et al.CVPR 2021
- LVFace: Progressive Cluster Optimization for Large Vision Models in Face RecognitionJinghan You, Shanglin Li, Yuanrui Sun, Jiangchuan Wei et al.ICCV 2025 · 3 citations
