Explicitly Minimizing the Blur Error of Variational Autoencoders
Gustav Bredell, Kyriakos Flouris, Krishna Chaitanya, Ertunc Erdil, Ender Konukoglu
Abstract
Variational autoencoders (VAEs) are powerful generative modelling methods, however they suffer from blurry generated samples and reconstructions compared to the images they have been trained on. Significant research effort has been spent to increase the generative capabilities by creating more flexible models but often flexibility comes at the cost of higher complexity and computational cost. Several works have focused on altering the reconstruction term of the evidence lower bound (ELBO), however, often at the expense of losing the mathematical link to maximizing the likelihood of the samples under the modeled distribution. Here we propose a new formulation of the reconstruction term for the VAE that specifically penalizes the generation of blurry images while at the same time still maximizing the ELBO under the modeled distribution. We show the potential of the proposed loss on three different data sets, where it outperforms several recently proposed reconstruction losses for VAEs.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext abb86a2b-edb9-4589-b895-89c6756902edCited by top-tier papers10
- Deep Generative Clustering with Multimodal Diffusion Variational AutoencodersEmanuele Palumbo, Laura Manduchi, Sonia Laguna, Daphné Chopard et al.ICLR 2024 · 21 citations
- Distributional Learning of Variational AutoEncoder: Application to Synthetic Data GenerationSeunghwan An, Jong-June JeonNeurIPS 2023 · 19 citations
- Tree Variational AutoencodersLaura Manduchi, Moritz Vandenhirtz, Alain Ryser, Julia E. VogtNeurIPS 2023 · 17 citations
- Diffusion Prior Interpolation for Flexibility Real-World Face Super-ResolutionJiarui Yang, Tao Dai, Yufei Zhu, Naiqi Li et al.AAAI 2025 · 11 citations
- FELLE: Autoregressive Speech Synthesis with Token-Wise Coarse-to-Fine Flow MatchingHui Wang, Shujie Liu, Lingwei Meng, Jinyu Li et al.ACM MM 2025 · 1 citation
Builds on8
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- NVAE: A Deep Hierarchical Variational AutoencoderArash Vahdat, Jan KautzNeurIPS 2020 · 1,141 citations
- Stochastic Latent Actor-Critic: Deep Reinforcement Learning with a Latent Variable ModelAlex X. Lee, Anusha Nagabandi, Pieter Abbeel, Sergey LevineNeurIPS 2020 · 437 citations
- Focal Frequency Loss for Image Reconstruction and SynthesisLiming Jiang, Bo Dai, Wayne Wu, Chen Change LoyICCV 2021 · 422 citations
- Flows for simultaneous manifold learning and density estimationJohann Brehmer, Kyle CranmerNeurIPS 2020 · 187 citations
Related papers
- A Loss Function for Generative Neural Networks Based on Watson's Perceptual ModelSteffen Czolbe, Oswin Krause, Ingemar J. Cox, Christian IgelNeurIPS 2020 · 71 citations
- Generalization Gap in Amortized InferenceMingtian Zhang, Peter Hayes, David BarberNeurIPS 2022 · 14 citations
- Dual Contradistinctive Generative AutoencoderGaurav Parmar, Dacheng Li, Kwonjoon Lee, Zhuowen TuCVPR 2021
- From Variational to Deterministic AutoencodersPartha Ghosh, Mehdi S. M. Sajjadi, Antonio Vergari, Michael J. Black et al.ICLR 2020 · 298 citations
- Statistical Guarantees for Variational Autoencoders using PAC-Bayesian TheorySokhna Diarra Mbacke, Florence Clerc, Pascal GermainNeurIPS 2023 · 22 citations
