Function-space Parameterization of Neural Networks for Sequential Learning
Aidan Scannell, Riccardo Mereu, Paul Edmund Chang, Ella Tamir, Joni Pajarinen, Arno Solin
Abstract
Sequential learning paradigms pose challenges for gradient-based deep learning due to difficulties incorporating new data and retaining prior knowledge. While Gaussian processes elegantly tackle these problems, they struggle with scalability and handling rich inputs, such as images. To address these issues, we introduce a technique that converts neural networks from weight space to function space, through a dual parameterization. Our parameterization offers: (i) a way to scale function-space methods to large data sets via sparsification, (ii) retention of prior knowledge when access to past data is limited, and (iii) a mechanism to incorporate new data without retraining. Our experiments demonstrate that we can retain knowledge in continual learning and incorporate new data efficiently. We further show its strengths in uncertainty quantification and guiding exploration in model-based RL. Further information and code is available on the project website 1 .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers6
- FSP-Laplace: Function-Space Priors for the Laplace Approximation in Bayesian Deep LearningTristan Cinquin, Marvin Pförtner, Vincent Fortuin, Philipp Hennig et al.NeurIPS 2024 · 15 citations
- Post-hoc Probabilistic Vision-Language ModelsAnton Baumann, Rui Li, Marcus Klasson, Santeri Mentu et al.ICLR 2026 · 14 citations
- Variational Linearized Laplace Approximation for Bayesian Deep LearningLuis A. Ortega Andrés, Simón Rodríguez Santana, Daniel Hernández-LobatoICML 2024 · 12 citations
- Compact Memory for Continual Logistic RegressionYohan Jung, Hyungi Lee, Wenlong Chen, Thomas Möllenhoff et al.NeurIPS 2025 · 2 citations
- Discrete Codebook World Models for Continuous ControlAidan Scannell, Mohammadreza Nakhaeinezhadfard, Kalle Kujanpää, Yi Zhao et al.ICLR 2025
Builds on13
- Dark Experience for General Continual Learning: a Strong, Simple BaselinePietro Buzzega, Matteo Boschini, Angelo Porrello, Davide Abati et al.NeurIPS 2020 · 1,494 citations
- Laplace Redux - Effortless Bayesian Deep LearningErik A. Daxberger, Agustinus Kristiadi, Alexander Immer, Runa Eschenhagen et al.NeurIPS 2021 · 508 citations
- Functional Regularisation for Continual Learning with Gaussian ProcessesMichalis K. Titsias, Jonathan Schwarz, Alexander G. de G. Matthews, Razvan Pascanu et al.ICLR 2020 · 209 citations
- Continual Deep Learning by Functional Regularisation of Memorable PastPingbo Pan, Siddharth Swaroop, Alexander Immer, Runa Eschenhagen et al.NeurIPS 2020 · 179 citations
- Scalable Marginal Likelihood Estimation for Model Selection in Deep LearningAlexander Immer, Matthias Bauer, Vincent Fortuin, Gunnar Rätsch et al.ICML 2021 · 130 citations
Related papers
- Continual Learning via Sequential Function-Space Variational InferenceTim G. J. Rudner, Freddie Bickford Smith, Qixuan Feng, Yee Whye Teh et al.ICML 2022 · 57 citations
- Memory-Based Dual Gaussian Processes for Sequential LearningPaul Edmund Chang, Prakhar Verma, S. T. John, Arno Solin et al.ICML 2023 · 10 citations
- Growing a Brain with Sparsity-Inducing Generation for Continual LearningHyundong Jin, Gyeong-Hyeon Kim, Chanho Ahn, Eunwoo KimICCV 2023 · 7 citations
- Deep Neural Networks as Point Estimates for Deep Gaussian ProcessesVincent Dutordoir, James Hensman, Mark van der Wilk, Carl Henrik Ek et al.NeurIPS 2021 · 35 citations
- Activation-level uncertainty in deep neural networksPablo Morales-Alvarez, Daniel Hernández-Lobato, Rafael Molina, José Miguel Hernández-LobatoICLR 2021 · 16 citations
