Quantum-Gated Task-interaction Knowledge Distillation for Pre-trained Model-based Class-Incremental Learning
Linjie Li, HUIYU XIAO, Jiarui Cao, Zhenyu Wu, Yang Ji
摘要
Class-incremental learning (CIL) aims to continuously accumulate knowledge from a stream of tasks and construct a unified classifier over all previously seen classes. A key challenge of CIL lies in the discrepancy between clear task boundaries during training and blurred boundaries during inference, where samples from different tasks often occupy overlapping subspaces. Although pretrained models (PTMs) have shown promising performance in CIL, they still struggle with the entanglement of multi-task subspaces, leading to catastrophic forgetting when task routing parameters are poorly calibrated or task-level representations are rigidly fixed. To address this issue, we propose a novel Quantum-Gated Task-interaction Knowledge Distillation (QKD) framework that leverages quantum gating to guide inter-task knowledge transfer. Specifically, we introduce a quantum-gated task modulation gating mechanism to model the relational dependencies among task embedding, dynamically capturing the sample-to-task relevance for both joint training and inference across streaming tasks. Furthermore, we employ lightweight adapters to adapt PTMs to downstream tasks while freezing previously learned adapters. Guided by the quantum gating outputs, we perform task-interaction knowledge distillation guided by these task-embedding-level correlation weights from old to new adapters, enabling the model to bridge the representation gaps between independent task subspaces and jointly calibrate the unified classifier. Extensive experiments on five benchmark datasets demonstrate that QKD effectively mitigates catastrophic forgetting and achieves state-of-the-art performance in class-incremental settings.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper16
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- The Many Faces of Robustness: A Critical Analysis of Out-of-Distribution GeneralizationDan Hendrycks, Steven Basart, Norman Mu, Saurav Kadavath 等ICCV 2021 · 被引用 2,294 次
- AdaptFormer: Adapting Vision Transformers for Scalable Visual RecognitionShoufa Chen, Chongjian Ge, Zhan Tong, Jiangliu Wang 等NeurIPS 2022 · 被引用 1,291 次
- Learning to Prompt for Continual LearningZifeng Wang, Zizhao Zhang, Chen-Yu Lee, Han Zhang 等CVPR 2022 · 被引用 635 次
- Scaling & Shifting Your Features: A New Baseline for Efficient Model TuningDongze Lian, Daquan Zhou, Jiashi Feng, Xinchao WangNeurIPS 2022 · 被引用 415 次
相关 Paper
- Parameter-Masked Decoupled Optimization for Cross-Domain Class-Incremental LearningZiqi Gu, Chunyan Xu, Yangguang Liu, Wenxuan Fang 等ICML 2026
- Maintaining Fairness in Logit-based Knowledge Distillation for Class-Incremental LearningZijian Gao, Shanhao Han, Xingxing Zhang, Kele Xu 等AAAI 2025 · 被引用 10 次
- MOS: Model Surgery for Pre-Trained Model-Based Class-Incremental LearningHai-Long Sun, Da-Wei Zhou, Hanbin Zhao, Le Gan 等AAAI 2025 · 被引用 31 次
- Compress to One Point: Neural Collapse for Pre-Trained Model-Based Class-Incremental LearningKun Wei, Zhe Xu, Cheng DengAAAI 2025 · 被引用 3 次
- Federated Continual Learning via Orchestrating Multi-Scale ExpertiseXiaoyang Yi, Yang Liu, Binhan Yang, Jian Jun ZhangNeurIPS 2025
