All in the Exponential Family: Bregman Duality in Thermodynamic Variational Inference
Rob Brekelmans, Vaden Masrani, Frank Wood, Greg Ver Steeg, Aram Galstyan
Abstract
The recently proposed Thermodynamic Variational Objective (TVO) leverages thermodynamic integration to provide a family of variational inference objectives, which both tighten and generalize the ubiquitous Evidence Lower Bound (ELBO). However, the tightness of TVO bounds was not previously known, an expensive grid search was used to choose a "schedule" of intermediate distributions, and model learning suffered with ostensibly tighter bounds. In this work, we propose an exponential family interpretation of the geometric mixture curve underlying the TVO and various path sampling methods, which allows us to characterize the gap in TVO likelihood bounds as a sum of KL divergences. We propose to choose intermediate distributions using equal spacing in the moment parameters of our exponential family, which matches grid search performance and allows the schedule to adaptively update over the course of training. Finally, we derive a doubly reparameterized gradient estimator which improves model learning and allows the TVO to benefit from more refined bounds. To further contextualize our contributions, we provide a unified framework for understanding thermodynamic integration and the TVO using Taylor series remainders.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a7f79dbf-0de1-4f22-893f-15f80cb5409fCited by top-tier papers9
- Trust Region Constrained Measure Transport in Path Space for Stochastic Optimal Control and InferenceDenis Blessing, Julius Berner, Lorenz Richter, Carles Domingo-Enrich et al.NeurIPS 2025 · 24 citations
- A connection between Tempering and Entropic Mirror DescentNicolas Chopin, Francesca R. Crucinio, Anna KorbaICML 2024 · 22 citations
- Provable benefits of annealing for estimating normalizing constants: Importance Sampling, Noise-Contrastive Estimation, and beyondOmar Chehab, Aapo Hyvärinen, Andrej RisteskiNeurIPS 2023 · 18 citations
- Hamiltonian Dynamics with Non-Newtonian Momentum for Rapid SamplingGreg Ver Steeg, Aram GalstyanNeurIPS 2021 · 18 citations
- ItDPDM: Information-Theoretic Discrete Poisson Diffusion ModelSagnik Bhattacharya, Abhiram Rao Gorle, Ahsan Bilal, Connor Ding et al.NeurIPS 2025 · 6 citations
Builds on1
Related papers
- Gaussian Process Bandit Optimization of the Thermodynamic Variational ObjectiveVu Nguyen, Vaden Masrani, Rob Brekelmans, Michael A. Osborne et al.NeurIPS 2020 · 5 citations
- Optimal Variance Control of the Score-Function Gradient Estimator for Importance-Weighted BoundsValentin Liévin, Andrea Dittadi, Anders Christensen, Ole WintherNeurIPS 2020 · 9 citations
- Nested Variational InferenceHeiko Zimmermann, Hao Wu, Babak Esmaeili, Jan-Willem van de MeentNeurIPS 2021 · 26 citations
- Variational Inference with Locally Enhanced Bounds for Hierarchical ModelsTomas Geffner, Justin DomkeICML 2022 · 6 citations
- VarGrad: A Low-Variance Gradient Estimator for Variational InferenceLorenz Richter, Ayman Boustati, Nikolas Nüsken, Francisco J. R. Ruiz et al.NeurIPS 2020 · 90 citations
