Embracing the Dark Knowledge: Domain Generalization Using Regularized Knowledge Distillation
Yufei Wang, Haoliang Li, Lap-Pui Chau, Alex C. Kot
Abstract
Though convolutional neural networks are widely used in different tasks, lack of generalization capability in the absence of sufficient and representative data is one of the challenges that hinders their practical application. In this paper, we propose a simple, effective, and plug-and-play training strategy named Knowledge Distillation for Domain Generalization (KDDG) which is built upon a knowledge distillation framework with the gradient filter as a novel regularization term. We find that both the "richer dark knowledge" from the teacher network, as well as the gradient filter we proposed, can reduce the difficulty of learning the mapping which further improves the generalization ability of the model. We also conduct experiments extensively to show that our framework can significantly improve the generalization capability of deep neural networks in different tasks including image classification, segmentation, reinforcement learning by comparing our method with existing state-of-the-art domain generalization techniques. Last but not the least, we propose to adopt two metrics to analyze our proposed method in order to better understand how our proposed method benefits the generalization capability of deep neural networks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers13
- Part-Aware Transformer for Generalizable Person Re-identificationHao Ni, Yuke Li, Lianli Gao, Heng Tao Shen et al.ICCV 2023 · 87 citations
- RDA: Robust Domain Adaptation via Fourier Adversarial AttackingJiaxing Huang, Dayan Guan, Aoran Xiao, Shijian LuICCV 2021 · 85 citations
- A Sentence Speaks a Thousand Images: Domain Generalization through Distilling CLIP with Language GuidanceZeyi Huang, Andy Zhou, Zijian Lin, Mu Cai et al.ICCV 2023 · 56 citations
- Label-Efficient Domain Generalization via Collaborative Exploration and GeneralizationJunkun Yuan, Xu Ma, Defang Chen, Kun Kuang et al.ACM MM 2022 · 21 citations
- Bayesian Knowledge Distillation: A Bayesian Perspective of Distillation with Uncertainty QuantificationLuyang Fang, Yongkai Chen, Wenxuan Zhong, Ping MaICML 2024 · 10 citations
Builds on6
- Moment Matching for Multi-Source Domain AdaptationXingchao Peng, Qinxun Bai, Xide Xia, Zijun Huang et al.ICCV 2019 · 2,239 citations
- Deep Domain-Adversarial Image Generation for Domain GeneralisationKaiyang Zhou, Yongxin Yang, Timothy M. Hospedales, Tao XiangAAAI 2020 · 488 citations
- Episodic Training for Domain GeneralizationDa Li, Jianshu Zhang, Yongxin Yang, Cong Liu et al.ICCV 2019 · 488 citations
- Invariant Risk Minimization GamesKartik Ahuja, Karthikeyan Shanmugam, Kush R. Varshney, Amit DhurandharICML 2020 · 289 citations
- Domain Generalization for Medical Imaging Classification with Linear-Dependency RegularizationHaoliang Li, Yufei Wang, Renjie Wan, Shiqi Wang et al.NeurIPS 2020 · 233 citations
Related papers
- Adversarial Teacher-Student Representation Learning for Domain GeneralizationFu-En Yang, Yuan-Chia Cheng, Zu-Yun Shiau, Yu-Chiang Frank WangNeurIPS 2021 · 83 citations
- FreeKD: Free-direction Knowledge Distillation for Graph Neural NetworksKaituo Feng, Changsheng Li, Ye Yuan, Guoren WangKDD 2022 · 28 citations
- Be Your Own Teacher: Improve the Performance of Convolutional Neural Networks via Self DistillationLinfeng Zhang, Jiebo Song, Anni Gao, Jingwei Chen et al.ICCV 2019 · 1,069 citations
- Revisiting Knowledge Distillation via Label Smoothing RegularizationLi Yuan, Francis E. H. Tay, Guilin Li, Tao Wang et al.CVPR 2020
- Revisiting Knowledge Distillation: An Inheritance and Exploration FrameworkZhen Huang, Xu Shen, Jun Xing, Tongliang Liu et al.CVPR 2021
