Posterior Meta-Replay for Continual Learning
Christian Henning, Maria R. Cervera, Francesco D'Angelo, Johannes von Oswald, Regina Traber, Benjamin Ehret, Seijin Kobayashi, Benjamin F. Grewe, João Sacramento
摘要
Continual Learning (CL) algorithms have recently received a lot of attention as they attempt to overcome the need to train with an i.i.d. sample from some unknown target data distribution. Building on prior work, we study principled ways to tackle the CL problem by adopting a Bayesian perspective and focus on continually learning a task-specific posterior distribution via a shared meta-model, a task-conditioned hypernetwork. This approach, which we term Posterior-replay CL, is in sharp contrast to most Bayesian CL approaches that focus on the recursive update of a single posterior distribution. The benefits of our approach are (1) an increased flexibility to model solutions in weight space and therewith less susceptibility to task dissimilarity, (2) access to principled task-specific predictive uncertainty estimates, that can be used to infer task identity during test time and to detect task boundaries during training, and (3) the ability to revisit and update task-specific posteriors in a principled manner without requiring access to past data. The proposed framework is versatile, which we demonstrate using simple posterior approximations (such as Gaussians) as well as powerful, implicit distributions modelled via a neural network. We illustrate the conceptual advance of our framework on low-dimensional problems and show performance gains on computer vision benchmarks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper22
- Repulsive Deep Ensembles are BayesianFrancesco D'Angelo, Vincent FortuinNeurIPS 2021 · 被引用 141 次
- A Theoretical Study on Solving Continual LearningGyuhak Kim, Changnan Xiao, Tatsuya Konishi, Zixuan Ke 等NeurIPS 2022 · 被引用 119 次
- A Unified and General Framework for Continual LearningZhenyi Wang, Yan Li, Li Shen, Heng HuangICLR 2024 · 被引用 42 次
- Learnability and Algorithm for Continual LearningGyuhak Kim, Changnan Xiao, Tatsuya Konishi, Bing LiuICML 2023 · 被引用 37 次
- Replay-and-Forget-Free Graph Class-Incremental Learning: A Task Profiling and Prompting ApproachChaoxi Niu, Guansong Pang, Ling Chen, Bing LiuNeurIPS 2024 · 被引用 32 次
它引用的顶会 Paper21
- Continual learning in recurrent neural networksBenjamin Ehret, Christian Henning, Maria R. Cervera, Alexander Meulemans 等ICLR 2021 · 被引用 4,433 次
- Bayesian Deep Learning and a Probabilistic Perspective of GeneralizationAndrew Gordon Wilson, Pavel IzmailovNeurIPS 2020 · 被引用 845 次
- Simple and Principled Uncertainty Estimation with Deterministic Deep Learning via Distance AwarenessJeremiah Z. Liu, Zi Lin, Shreyas Padhy, Dustin Tran 等NeurIPS 2020 · 被引用 604 次
- Continual learning with hypernetworksJohannes von Oswald, Christian Henning, João Sacramento, Benjamin F. GreweICLR 2020 · 被引用 412 次
- How Good is the Bayes Posterior in Deep Neural Networks Really?Florian Wenzel, Kevin Roth, Bastiaan S. Veeling, Jakub Swiatkowski 等ICML 2020 · 被引用 409 次
相关 Paper
- Uncertainty-guided Continual Learning with Bayesian Neural NetworksSayna Ebrahimi, Mohamed Elhoseiny, Trevor Darrell, Marcus RohrbachICLR 2020 · 被引用 211 次
- Kalman Filter for Online Classification of Non-Stationary DataMichalis K. Titsias, Alexandre Galashov, Amal Rannen-Triki, Razvan Pascanu 等ICLR 2024 · 被引用 14 次
- Temporal-Difference Variational Continual LearningLuckeciano Carvalho Melo, Alessandro Abate, Yarin GalNeurIPS 2025 · 被引用 1 次
- Functional Regularisation for Continual Learning with Gaussian ProcessesMichalis K. Titsias, Jonathan Schwarz, Alexander G. de G. Matthews, Razvan Pascanu 等ICLR 2020 · 被引用 209 次
- VariGrow: Variational Architecture Growing for Task-Agnostic Continual Learning based on Bayesian NoveltyRandy Ardywibowo, Zepeng Huo, Zhangyang Wang, Bobak J. Mortazavi 等ICML 2022 · 被引用 11 次
