Meta-Consolidation for Continual Learning
K. J. Joseph, Vineeth Nallure Balasubramanian
Abstract
The ability to continuously learn and adapt itself to new tasks, without losing grasp of already acquired knowledge is a hallmark of biological learning systems, which current deep learning systems fall short of. In this work, we present a novel methodology for continual learning called MERLIN: Meta-Consolidation for Continual Learning. We assume that weights of a neural network , for solving task , come from a meta-distribution . This meta-distribution is learned and consolidated incrementally. We operate in the challenging online continual learning setting, where a data point is seen by the model only once. Our experiments with continual learning benchmarks of MNIST, CIFAR-10, CIFAR-100 and Mini-ImageNet datasets show consistent improvement over five baselines, including a recent state-of-the-art, corroborating the promise of MERLIN.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8a4de93a-2559-4742-83c5-f00a492f4809Cited by top-tier papers10
- Posterior Meta-Replay for Continual LearningChristian Henning, Maria R. Cervera, Francesco D'Angelo, Johannes von Oswald et al.NeurIPS 2021 · 78 citations
- Optimizing Reusable Knowledge for Continual Learning via MetalearningJulio Hurtado, Alain Raymond-Saez, Alvaro SotoNeurIPS 2021 · 47 citations
- Formalizing the Generalization-Forgetting Trade-off in Continual LearningKrishnan Raghavan, Prasanna BalaprakashNeurIPS 2021 · 42 citations
- Energy-based Latent Aligner for Incremental LearningK. J. Joseph, Salman Khan, Fahad Shahbaz Khan, Rao Muhammad Anwer et al.CVPR 2022 · 35 citations
- Objects in Semantic TopologyShuo Yang, Peize Sun, Yi Jiang, Xiaobo Xia et al.ICLR 2022 · 35 citations
Builds on6
- Continual learning with hypernetworksJohannes von Oswald, Christian Henning, João Sacramento, Benjamin F. GreweICLR 2020 · 412 citations
- Task2Vec: Task Embedding for Meta-LearningAlessandro Achille, Michael Lam, Rahul Tewari, Avinash Ravichandran et al.ICCV 2019 · 359 citations
- A Neural Dirichlet Process Mixture Model for Task-Free Continual LearningSoochan Lee, Junsoo Ha, Dongsu Zhang, Gunhee KimICLR 2020 · 238 citations
- Uncertainty-guided Continual Learning with Bayesian Neural NetworksSayna Ebrahimi, Mohamed Elhoseiny, Trevor Darrell, Marcus RohrbachICLR 2020 · 211 citations
- Functional Regularisation for Continual Learning with Gaussian ProcessesMichalis K. Titsias, Jonathan Schwarz, Alexander G. de G. Matthews, Razvan Pascanu et al.ICLR 2020 · 209 citations
Related papers
- Bayesian Structural Adaptation for Continual LearningAbhishek Kumar, Sunabha Chatterjee, Piyush RaiICML 2021 · 7 citations
- Learning to Continually Learn with the Bayesian PrincipleSoochan Lee, Hyeonseong Jeon, Jaehyeon Son, Gunhee KimICML 2024 · 11 citations
- Batch Model Consolidation: A Multi-Task Model Consolidation FrameworkIordanis Fostiropoulos, Jiaye Zhu, Laurent IttiCVPR 2023
- iTAML: An Incremental Task-Agnostic Meta-learning ApproachJathushan Rajasegaran, Salman H. Khan, Munawar Hayat, Fahad Shahbaz Khan et al.CVPR 2020
- Mitigating Forgetting in Online Continual Learning via Instance-Aware ParameterizationHung-Jen Chen, An-Chieh Cheng, Da-Cheng Juan, Wei Wei et al.NeurIPS 2020 · 50 citations
