Bayesian Adaptation of Network Depth and Width for Continual Learning
Jeevan Thapa, Rui Li
摘要
While existing dynamic architecture-based continual learning methods adapt network width by growing new branches, they overlook the critical aspect of network depth. We propose a novel non-parametric Bayesian approach to infer network depth and adapt network width while maintaining model performance across tasks. Specifically, we model the growth of network depth with a beta process and apply drop-connect regularization to network width using a conjugate Bernoulli process. Our results show that our proposed method achieves superior or comparable performance with state-of-the-art methods across various continual learning benchmarks. Moreover, our approach can be readily extended to unsupervised continual learning, showcasing competitive performance compared to existing techniques.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Temporal-Difference Variational Continual LearningLuckeciano Carvalho Melo, Alessandro Abate, Yarin GalNeurIPS 2025 · 被引用 1 次
- Multi-Synaptic Cooperation: A Bio-Inspired Framework for Robust and Scalable Continual LearningPenghui Li, Zhuang Ma, Yunliang Zang, Qiang YuICLR 2026
它引用的顶会 Paper12
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- New Insights on Reducing Abrupt Representation Change in Online Continual LearningLucas Caccia, Rahaf Aljundi, Nader Asadi, Tinne Tuytelaars 等ICLR 2022 · 被引用 279 次
- Uncertainty-guided Continual Learning with Bayesian Neural NetworksSayna Ebrahimi, Mohamed Elhoseiny, Trevor Darrell, Marcus RohrbachICLR 2020 · 被引用 211 次
- Lifelong GAN: Continual Learning for Conditional Image GenerationMengyao Zhai, Lei Chen, Frederick Tung, Jiawei He 等ICCV 2019 · 被引用 204 次
- Depth Uncertainty in Neural NetworksJavier Antorán, James Urquhart Allingham, José Miguel Hernández-LobatoNeurIPS 2020 · 被引用 121 次
相关 Paper
- Joint Inference for Neural Network Depth and Dropout RegularizationKishan K. C., Rui Li, Mahdi GilanyNeurIPS 2021 · 被引用 13 次
- AdaVAE: Bayesian Structural Adaptation for Variational AutoencodersParibesh Regmi, Rui LiNeurIPS 2023 · 被引用 4 次
- Bayesian Structural Adaptation for Continual LearningAbhishek Kumar, Sunabha Chatterjee, Piyush RaiICML 2021 · 被引用 7 次
- Functional Regularisation for Continual Learning with Gaussian ProcessesMichalis K. Titsias, Jonathan Schwarz, Alexander G. de G. Matthews, Razvan Pascanu 等ICLR 2020 · 被引用 209 次
- VariGrow: Variational Architecture Growing for Task-Agnostic Continual Learning based on Bayesian NoveltyRandy Ardywibowo, Zepeng Huo, Zhangyang Wang, Bobak J. Mortazavi 等ICML 2022 · 被引用 11 次
