Probabilistic Group Mask Guided Discrete Optimization for Incremental Learning
Fengqiang Wan, Yang Yang
Abstract
Incremental learning (IL) aims to sequentially learn new tasks while mitigating catastrophic forgetting. Among various IL strategies, parameterisolation methods stand out by using mask techniques to allocate distinct parameters to each task, explicitly addressing forgetting. However, existing approaches often disregard parameter dependencies, resulting in an over-reliance on newly allocated parameters. To address this issue, we propose Probabilistic Group Mask selection (PGM), a group-wise approach that captures parameter dependencies by exploring candidate masks within each group. Specifically, PGM partitions parameters into groups with multiple candidate masks, assigning probabilities to these masks and leveraging Gumbel-Softmax for differentiable sampling, enabling efficient optimization of the discrete mask selection process. Our theoretical analysis demonstrates that incorporating parameter dependencies enhances subnetwork selection. Experiments conducted on standard benchmarks confirm its superior effectiveness compared to existing IL approaches. The source code is available at: https://github. com/njustkmg/ICML25-PGM .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b3caa36a-740a-44d9-a87f-49359f93d2d6Cited by top-tier papers2
- Shaping Parameter Contribution Patterns for Out-of-Distribution DetectionHaonan Xu, Yang YangAAAI 2026
- GR-LoRA: Gradient-Recycling Low-Rank Adaptation for Class-Incremental LearningYipeng Lin, Fengqiang Wan, Yang YangICML 2026
Builds on15
- Energy-based Out-of-distribution DetectionWeitang Liu, Xiaoyun Wang, John D. Owens, Yixuan LiNeurIPS 2020 · 2,213 citations
- Gradient Projection Memory for Continual LearningGobinda Saha, Isha Garg, Kaushik RoyICLR 2021 · 409 citations
- Supermasks in SuperpositionMitchell Wortsman, Vivek Ramanujan, Rosanne Liu, Aniruddha Kembhavi et al.NeurIPS 2020 · 364 citations
- Online Coreset Selection for Rehearsal-based Continual LearningJaehong Yoon, Divyam Madaan, Eunho Yang, Sung Ju HwangICLR 2022 · 181 citations
- Forget-free Continual Learning with Winning SubnetworksHaeyong Kang, Rusty John Lloyd Mina, Sultan Rizky Hikmawan Madjid, Jaehong Yoon et al.ICML 2022 · 159 citations
Related papers
- Parameter-Level Soft-Masking for Continual LearningTatsuya Konishi, Mori Kurokawa, Chihiro Ono, Zixuan Ke et al.ICML 2023 · 63 citations
- Order-Robust Class Incremental Learning: Graph-Driven Dynamic Similarity GroupingGuannan Lai, Yujie Li, Xiangkun Wang, Junbo Zhang et al.CVPR 2025
- Towards Robust Graph Incremental Learning on Evolving GraphsJunwei Su, Difan Zou, Zijun Zhang, Chuan WuICML 2023 · 37 citations
- Parameter-Masked Decoupled Optimization for Cross-Domain Class-Incremental LearningZiqi Gu, Chunyan Xu, Yangguang Liu, Wenxuan Fang et al.ICML 2026
- Learn or Recall? Revisiting Incremental Learning with Pre-trained Language ModelsJunhao Zheng, Shengjie Qiu, Qianli MaACL 2024
