Maintaining Fairness in Logit-based Knowledge Distillation for Class-Incremental Learning
Zijian Gao, Shanhao Han, Xingxing Zhang, Kele Xu, Dulan Zhou, Xinjun Mao, Yong Dou, Huaimin Wang
摘要
Logit-based knowledge distillation (KD) is commonly used to mitigate catastrophic forgetting in class-incremental learning (CIL) caused by data distribution shifts. However, the strict match of logit values between student and teacher models conflicts with the cross-entropy (CE) loss objective of learning new classes, leading to significant recency bias (i.e. unfairness). To address this issue, we rethink the overlooked limitations of KD-based methods through empirical analysis. Inspired by our findings, we introduce a plug-and-play preprocess method that normalizes the logits of both the student and teacher across all classes, rather than just the old classes, before distillation. This approach allows the student to focus on both old and new classes, capturing intrinsic inter-class relations from the teacher. By doing so, our method avoids the inherent conflict between KD and CE, maintaining fairness between old and new classes. Additionally, recognizing that overconfident teacher predictions can hinder the transfer of inter-class relations (i.e., dark knowledge), we extend our method to capture intra-class relations among different instances, ensuring fairness within old classes. Our method integrates seamlessly with existing logit-based KD approaches, consistently enhancing their performance across multiple CIL benchmarks without incurring additional training costs.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- Pansharpening for Thin-Cloud Contaminated Remote Sensing Images: A Unified Framework and Benchmark DatasetSongcheng Du, Yang Zou, Jiaxin Li, Mingxuan Liu 等AAAI 2026 · 被引用 3 次
- Unified Representation Causal Prompt Distillation for Re-Inference-Free Lifelong Person Re-IdentificationJiaqi Zhao, Jie Luo, Yong Zhou, Wen-Liang Du 等AAAI 2026
- Decouple Your Discovery and Memory in Continual Generalized Category DiscoveryJiawei Yu, Zijian Gao, Xingxing Zhang, Xuan Liu 等CVPR 2026
- Re-evaluating Continual VQA: Toward Fair and Robust Evaluation for Multimodal Continual LearningZijian Gao, Zicheng Sun, Xingxing Zhang, Kele Xu 等CVPR 2026
- Direction Sensitivity-Based Knowledge Distillation: Optimization-Aware Low-Rank Knowledge TransferYongkai Liao, Xinxing Chen, Zhongzheng Fu, Haoyuan Wang 等AAAI 2026
它引用的顶会 Paper8
- Similarity-Preserving Knowledge DistillationFrederick Tung, Greg MoriICCV 2019 · 被引用 1,214 次
- Knowledge Distillation from A Stronger TeacherTao Huang, Shan You, Fei Wang, Chen Qian 等NeurIPS 2022 · 被引用 477 次
- Logit Standardization in Knowledge DistillationShangquan Sun, Wenqi Ren, Jingzhi Li, Rui Wang 等CVPR 2024 · 被引用 183 次
- Self-Sustaining Representation Expansion for Non-Exemplar Class-Incremental LearningKai Zhu, Wei Zhai, Yang Cao, Jiebo Luo 等CVPR 2022 · 被引用 155 次
- Agree to Disagree: Adaptive Ensemble Knowledge Distillation in Gradient SpaceShangchen Du, Shan You, Xiaojie Li, Jianlong Wu 等NeurIPS 2020 · 被引用 144 次
相关 Paper
- Maintaining Discrimination and Fairness in Class Incremental LearningBowen Zhao, Xi Xiao, Guojun Gan, Bin Zhang 等CVPR 2020
- Gradient Reweighting: Towards Imbalanced Class-Incremental LearningJiangpeng HeCVPR 2024
- SS-IL: Separated Softmax for Incremental LearningHongjoon Ahn, Jihwan Kwak, Subin Lim, Hyeonsu Bang 等ICCV 2021 · 被引用 209 次
- Resolving Task Confusion in Dynamic Expansion Architectures for Class Incremental LearningBingchen Huang, Zhineng Chen, Peng Zhou, Jiayin Chen 等AAAI 2023 · 被引用 31 次
- Defying Imbalanced Forgetting in Class Incremental LearningShixiong Xu, Gaofeng Meng, Xing Nie, Bolin Ni 等AAAI 2024 · 被引用 8 次
