Simple and Effective VAE Training with Calibrated Decoders
Oleh Rybkin, Kostas Daniilidis, Sergey Levine
Abstract
Variational autoencoders (VAEs) provide an effective and simple method for modeling complex distributions. However, training VAEs often requires considerable hyperparameter tuning, and often utilizes a heuristic weight on the prior KL-divergence term. In this work, we study how the performance of VAEs can be improved while not requiring the use of this heuristic hyperparameter, by learning calibrated decoders that accurately model the decoding distribution. While in some sense it may seem obvious that calibrated decoders should perform better than uncalibrated decoders, much of the recent literature that employs VAEs uses uncalibrated Gaussian decoders with constant variance. We observe empirically that the naive way of learning variance in Gaussian decoders does not lead to good results. However, other calibrated decoders, such as discrete decoders or learning shared variance can substantially improve performance. To further improve results, we propose a simple but novel modification to the commonly used Gaussian decoder, which represents the prediction variance non-parametrically. We observe empirically that using the heuristic weight hyperparameter is not necessary with our method. We analyze the performance of various discrete and continuous decoders on a range of datasets and several single-image and sequential VAE models. Project website: this https URL
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0677cf5f-df50-4069-8e27-f404fda7a7e4Cited by top-tier papers22
- Diffusion Autoencoders: Toward a Meaningful and Decodable RepresentationKonpat Preechakul, Nattanat Chatthee, Suttisak Wizadwongsa, Supasorn SuwajanakornCVPR 2022 · 276 citations
- Long-Horizon Visual Planning with Goal-Conditioned Hierarchical PredictorsKarl Pertsch, Oleh Rybkin, Frederik Ebert, Shenghao Zhou et al.NeurIPS 2020 · 96 citations
- Meta-GMVAE: Mixture of Gaussian VAE for Unsupervised Meta-LearningDong Bok Lee, Dongchan Min, Seanie Lee, Sung Ju HwangICLR 2021 · 62 citations
- PoseFix: Correcting 3D Human Poses with Natural LanguageGinger Delmas, Philippe Weinzaepfel, Francesc Moreno-Noguer, Grégory RogezICCV 2023 · 49 citations
- Probabilistic Weather Forecasting with Hierarchical Graph Neural NetworksJoel Oskarsson, Tomas Landelius, Marc Peter Deisenroth, Fredrik LindstenNeurIPS 2024 · 44 citations
Builds on6
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Dream to Control: Learning Behaviors by Latent ImaginationDanijar Hafner, Timothy P. Lillicrap, Jimmy Ba, Mohammad NorouziICLR 2020 · 1,852 citations
- Stochastic Latent Actor-Critic: Deep Reinforcement Learning with a Latent Variable ModelAlex X. Lee, Anusha Nagabandi, Pieter Abbeel, Sergey LevineNeurIPS 2020 · 437 citations
- Skew-Fit: State-Covering Self-Supervised Reinforcement LearningVitchyr Pong, Murtaza Dalal, Steven Lin, Ashvin Nair et al.ICML 2020 · 303 citations
- From Variational to Deterministic AutoencodersPartha Ghosh, Mehdi S. M. Sajjadi, Antonio Vergari, Michael J. Black et al.ICLR 2020 · 298 citations
Related papers
- Shape your Space: A Gaussian Mixture Regularization Approach to Deterministic AutoencodersAmrutha Saseendran, Kathrin Skubch, Stefan Falkner, Margret KeuperNeurIPS 2021 · 13 citations
- Variational Learning of Fractional PosteriorsKian Ming A. Chai, Edwin V. BonillaICML 2025
- Vector Quantized Wasserstein Auto-EncoderLong Tung Vuong, Trung Le, He Zhao, Chuanxia Zheng et al.ICML 2023 · 24 citations
- Learning Optimal Priors for Task-Invariant Representations in Variational AutoencodersHiroshi Takahashi, Tomoharu Iwata, Atsutoshi Kumagai, Sekitoshi Kanai et al.KDD 2022 · 4 citations
- Multi-Rate VAE: Train Once, Get the Full Rate-Distortion CurveJuhan Bae, Michael R. Zhang, Michael Ruan, Eric Wang et al.ICLR 2023 · 3 citations
