Generalization Gap in Amortized Inference
Mingtian Zhang, Peter Hayes, David Barber
Abstract
The ability of likelihood-based probabilistic models to generalize to unseen data is central to many machine learning applications such as lossless compression. In this work, we study the generalization of a popular class of probabilistic model - the Variational Auto-Encoder (VAE). We discuss the two generalization gaps that affect VAEs and show that overfitting is usually dominated by amortized inference. Based on this observation, we propose a new training objective that improves the generalization of amortized inference. We demonstrate how our method can improve performance in the context of image modeling and lossless compression.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 59c7f35b-66ed-46d6-8db8-c7b642ba82b3Cited by top-tier papers3
- Computationally-Efficient Neural Image Compression with Shallow DecodersYibo Yang, Stephan MandtICCV 2023 · 43 citations
- Laplacian Autoencoders for Learning Stochastic RepresentationsMarco Miani, Frederik Warburg, Pablo Moreno-Muñoz, Nicki Skafte Detlefsen et al.NeurIPS 2022 · 17 citations
- Moment Matching Denoising Gibbs SamplingMingtian Zhang, Alex Hawkins-Hooker, Brooks Paige, David BarberNeurIPS 2023 · 8 citations
Builds on6
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Improving Inference for Neural Image CompressionYibo Yang, Robert Bamler, Stephan MandtNeurIPS 2020 · 151 citations
- HiLLoC: lossless image compression with hierarchical latent variable modelsJames Townsend, Thomas Bird, Julius Kunze, David BarberICLR 2020 · 60 citations
- On the Out-of-distribution Generalization of Probabilistic Image ModellingMingtian Zhang, Andi Zhang, Steven McDonaghNeurIPS 2021 · 51 citations
- Improving Lossless Compression Rates via Monte Carlo Bits-Back CodingYangjun Ruan, Karen Ullrich, Daniel Severo, James Townsend et al.ICML 2021 · 25 citations
Related papers
- Trading Information between Latents in Hierarchical Variational AutoencodersTim Z. Xiao, Robert BamlerICLR 2023
- Information-theoretic Generalization Analysis for VQ-VAEs: A Role of Latent VariablesFutoshi Futami, Masahiro FujisawaNeurIPS 2025 · 1 citation
- Improving Variational Autoencoders with Density Gap-based RegularizationJianfei Zhang, Jun Bai, Chenghua Lin, Yanmeng Wang et al.NeurIPS 2022 · 11 citations
- Information-Theoretic Generalization Bounds for VAEs: A Role of Encoder and Latent VariableFutoshi Futami, Masahiro FujisawaICML 2026
- Hierarchical Quantized AutoencodersWill Williams, Sam Ringer, Tom Ash, David MacLeod et al.NeurIPS 2020 · 90 citations
