Distilling Virtual Examples for Long-tailed Recognition
Yin-Yin He, Jianxin Wu, Xiu-Shen Wei
Abstract
We tackle the long-tailed visual recognition problem from the knowledge distillation perspective by proposing a Distill the Virtual Examples (DiVE) method. Specifically, by treating the predictions of a teacher model as virtual examples, we prove that distilling from these virtual examples is equivalent to label distribution learning under certain constraints. We show that when the virtual example distribution becomes flatter than the original input distribution, the under-represented tail classes will receive significant improvements, which is crucial in long-tailed recognition. The proposed DiVE method can explicitly tune the virtual example distribution to become flat. Extensive experiments on three benchmark datasets, including the large-scale iNaturalist ones, justify that the proposed DiVE method can significantly outperform state-of-the-art methods. Further-more, additional analyses and experiments verify the virtual example interpretation, and demonstrate the effectiveness of tailored designs in DiVE for long-tailed problems.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers42
- Long- Tailed Recognition via Weight BalancingShaden Alshammari, Yu-Xiong Wang, Deva Ramanan, Shu KongCVPR 2022 · 133 citations
- Nested Collaborative Learning for Long-Tailed Visual RecognitionJun Li, Zichang Tan, Jun Wan, Zhen Lei et al.CVPR 2022 · 95 citations
- BatchFormer: Learning to Explore Sample Relationships for Robust Representation LearningZhi Hou, Baosheng Yu, Dacheng TaoCVPR 2022 · 92 citations
- Long-Tail Learning with Foundation Model: Heavy Fine-Tuning HurtsJiang-Xin Shi, Tong Wei, Zhi Zhou, Jie-Jing Shao et al.ICML 2024 · 78 citations
- Introspective Distillation for Robust Question AnsweringYulei Niu, Hanwang ZhangNeurIPS 2021 · 74 citations
Builds on13
- Decoupling Representation and Classifier for Long-Tailed RecognitionBingyi Kang, Saining Xie, Marcus Rohrbach, Zhicheng Yan et al.ICLR 2020 · 1,496 citations
- Long-tail learning via logit adjustmentAditya Krishna Menon, Sadeep Jayasumana, Ankit Singh Rawat, Himanshu Jain et al.ICLR 2021 · 937 citations
- Balanced Meta-Softmax for Long-Tailed Visual RecognitionJiawei Ren, Cunjun Yu, Shunan Sheng, Xiao Ma et al.NeurIPS 2020 · 861 citations
- Long-Tailed Classification by Keeping the Good and Removing the Bad Momentum Causal EffectKaihua Tang, Jianqiang Huang, Hanwang ZhangNeurIPS 2020 · 533 citations
- Rethinking the Value of Labels for Improving Class-Imbalanced LearningYuzhe Yang, Zhi XuNeurIPS 2020 · 512 citations
Related papers
- Self Supervision to Distillation for Long-Tailed Visual RecognitionTianhao Li, Limin Wang, Gangshan WuICCV 2021 · 122 citations
- Distilling Balanced Knowledge from a Biased TeacherSeonghak KimCVPR 2026 · 1 citation
- DeiT-LT: Distillation Strikes Back for Vision Transformer Training on Long-Tailed DatasetsHarsh Rangwani, Pradipto Mondal, Mayank Mishra, Ashish Ramayee Asokan et al.CVPR 2024 · 12 citations
- Disentangling Label Distribution for Long-Tailed Visual RecognitionYoungkyu Hong, Seungju Han, Kwanghee Choi, Seokjun Seo et al.CVPR 2021
- MDCS: More Diverse Experts with Consistency Self-distillation for Long-tailed RecognitionQihao Zhao, Chen Jiang, Wei Hu, Fan Zhang et al.ICCV 2023 · 35 citations
