Bayesian Online Natural Gradient (BONG)
Matt Jones, Peter G. Chang, Kevin P. Murphy
Abstract
We propose a novel approach to sequential Bayesian inference based on variational Bayes (VB). The key insight is that, in the online setting, we do not need to add the KL term to regularize to the prior (which comes from the posterior at the previous timestep); instead we can optimize just the expected log-likelihood, performing a single step of natural gradient descent starting at the prior predictive. We prove this method recovers exact Bayesian inference if the model is conjugate. We also show how to compute an efficient deterministic approximation to the VB objective, as well as our simplified objective, when the variational distribution is Gaussian or a sub-family, including the case of a diagonal plus low-rank precision matrix. We show empirically that our method outperforms other online VB methods in the non-conjugate setting, such as online learning for neural networks, especially when controlling for computational costs.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext beb7d4f1-93dc-4b19-9deb-b874ef8211ddCited by top-tier papers7
- Non-Stationary Learning of Neural Networks with Automatic Soft Parameter ResetAlexandre Galashov, Michalis K. Titsias, András György, Clare Lyle et al.NeurIPS 2024 · 18 citations
- Brain-like Variational InferenceHadi Vafaii, Dekel Galor, Jacob L. YatesNeurIPS 2025 · 7 citations
- SING: SDE Inference via Natural GradientsAmber Hu, Henry Smith, Scott W. LindermanNeurIPS 2025 · 6 citations
- Martingale Posterior Neural Networks for Fast Sequential Decision MakingGerardo Duran-Martin, Leandro Sánchez-Betancourt, Álvaro Cartea, Kevin MurphyNeurIPS 2025 · 5 citations
- Natural Gradient VI: Guarantees for Non-Conjugate ModelsFangyuan Sun, Ilyas Fatkhullin, Niao HeNeurIPS 2025 · 3 citations
Builds on9
- ADAHESSIAN: An Adaptive Second Order Optimizer for Machine LearningZhewei Yao, Amir Gholami, Sheng Shen, Mustafa Mustafa et al.AAAI 2021 · 358 citations
- Continual Learning with Bayesian Neural Networks for Non-Stationary DataRichard Kurle, Botond Cseke, Alexej Klushyn, Patrick van der Smagt et al.ICLR 2020 · 82 citations
- Variational Learning is Effective for Large Deep NetworksYuesong Shen, Nico Daheim, Bai Cong, Peter Nickl et al.ICML 2024 · 53 citations
- Efficient Low Rank Gaussian Variational Inference for Neural NetworksMarcin Tomczak, Siddharth Swaroop, Richard E. TurnerNeurIPS 2020 · 37 citations
- Outlier-robust Kalman Filtering through Generalised BayesGerardo Duran-Martin, Matías Altamirano, Alexander Y. Shestopaloff, Leandro Sánchez-Betancourt et al.ICML 2024 · 30 citations
Related papers
- Online Variational Filtering and Parameter LearningAndrew Campbell, Yuyang Shi, Thomas Rainforth, Arnaud DoucetNeurIPS 2021 · 30 citations
- Variational Auto-Regressive Gaussian Processes for Continual LearningSanyam Kapoor, Theofanis Karaletsos, Thang D. BuiICML 2021 · 32 citations
- Spike and slab variational Bayes for high dimensional logistic regressionKolyan Ray, Botond Szabó, Gabriel ClaraNeurIPS 2020 · 35 citations
- Forward-Backward Gaussian Variational Inference via JKO in the Bures-Wasserstein SpaceMichael Ziyang Diao, Krishna Balasubramanian, Sinho Chewi, Adil SalimICML 2023 · 47 citations
- Continual Learning via Sequential Function-Space Variational InferenceTim G. J. Rudner, Freddie Bickford Smith, Qixuan Feng, Yee Whye Teh et al.ICML 2022 · 57 citations
