Evaluation-oriented Knowledge Distillation for Deep Face Recognition
Yuge Huang, Jiaxiang Wu, Xingkun Xu, Shouhong Ding
摘要
Knowledge distillation (KD) is a widely-used technique that utilizes large networks to improve the performance of compact models. Previous KD approaches usually aim to guide the student to mimic the teacher's behavior completely in the representation space. However, such one-to-one corresponding constraints may lead to inflexible knowledge transfer from the teacher to the student, especially those with low model capacities. Inspired by the ultimate goal of KD methods, we propose a novel Evaluation-oriented KD method (EKD) for deep face recognition to directly reduce the performance gap between the teacher and student models during training. Specifically, we adopt the commonly used evaluation metrics in face recognition, i.e., False Positive Rate (FPR) and True Positive Rate (TPR) as the performance indicator. According to the evaluation protocol, the critical pair relations that cause the TPR and FPR difference between the teacher and student models are selected. Then, the critical relations in the student are constrained to approximate the corresponding ones in the teacher by a novel rank-based loss function, giving more flexibility to the student with low capacity. Extensive experimental results on popular benchmarks demonstrate the superiority of our EKD over state-of-the-art competitors.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Grouped Knowledge Distillation for Deep Face RecognitionWeisong Zhao, Xiangyu Zhu, Kaiwen Guo, Xiaoyu Zhang 等AAAI 2023 · 被引用 12 次
- Cross-View Consistency Regularisation for Knowledge DistillationWeijia Zhang, Dongnan Liu, Weidong Cai, Chao MaACM MM 2024 · 被引用 11 次
- Cross-Architecture Distillation for Face RecognitionWeisong Zhao, Xiangyu Zhu, Zhixiang He, Xiaoyu Zhang 等ACM MM 2023 · 被引用 10 次
- Foreground Object Search by Distilling Composite Image FeatureBo Zhang, Jiacheng Sui, Li NiuICCV 2023 · 被引用 8 次
- Cross-Architecture Distillation Made Simple with Redundancy SuppressionWeijia Zhang, Yuehao Liu, Wu Ran, Chao MaICCV 2025 · 被引用 6 次
它引用的顶会 Paper3
- Similarity-Preserving Knowledge DistillationFrederick Tung, Greg MoriICCV 2019 · 被引用 1,214 次
- Correlation Congruence for Knowledge DistillationBaoyun Peng, Xiao Jin, Dongsheng Li, Shunfeng Zhou 等ICCV 2019 · 被引用 625 次
- CurricularFace: Adaptive Curriculum Learning Loss for Deep Face RecognitionYuge Huang, Yuhan Wang, Ying Tai, Xiaoming Liu 等CVPR 2020
相关 Paper
- ICD-Face: Intra-class Compactness Distillation for Face RecognitionZhipeng Yu, Jiaheng Liu, Haoyu Qin, Yichao Wu 等ICCV 2023 · 被引用 7 次
- Revisit the Essence of Distilling Knowledge through CalibrationWen-Shu Fan, Su Lu, Xin-Chun Li, De-Chuan Zhan 等ICML 2024 · 被引用 8 次
- A Good Teacher Adapts Their Knowledge for DistillationChengyao Qian, Trung Le, Mehrtash HarandiICCV 2025 · 被引用 8 次
- Rethinking the Dark Knowledge and Kullback-Leibler Divergence Loss in Knowledge Distillation Under Capacity MismatchingYingchao Wang, Wenqi Niu, Xingshan Yao, Li You 等AAAI 2026
- Adaptive Dual Guidance Knowledge DistillationTong Li, Long Liu, Kang Liu, Xin Wang 等AAAI 2025 · 被引用 1 次
