Soft-IntroVAE: Analyzing and Improving the Introspective Variational Autoencoder
Tal Daniel, Aviv Tamar
Abstract
The recently introduced introspective variational autoencoder (IntroVAE) exhibits outstanding image generations, and allows for amortized inference using an image encoder. The main idea in IntroVAE is to train a VAE adversarially, using the VAE encoder to discriminate between generated and real data samples. However, the original In-troVAE loss function relied on a particular hinge-loss formulation that is very hard to stabilize in practice, and its theoretical convergence analysis ignored important terms in the loss. In this work, we take a step towards better understanding of the IntroVAE model, its practical implementation, and its applications. We propose the Soft-IntroVAE, a modified IntroVAE that replaces the hinge-loss terms with a smooth exponential loss on generated samples. This change significantly improves training stability, and also enables theoretical analysis of the complete algorithm. Interestingly, we show that the IntroVAE converges to a distribution that minimizes a sum of KL distance from the data distribution and an entropy term. We discuss the implications of this result, and demonstrate that it induces competitive image generation and reconstruction. Finally, we describe an application of Soft-IntroVAE to unsupervised image translation, and demonstrate compelling results. Code and additional information is available on the project websitetaldatech.github.io/soft-intro-vae-web.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers4
- Unsupervised Image Representation Learning with Deep Latent ParticlesTal Daniel, Aviv TamarICML 2022 · 18 citations
- Property Existence Inference against Generative ModelsLijin Wang, Jingjing Wang, Jie Wan, Lin Long et al.USENIX Security 2024 · 13 citations
- CODA: A Correlation-Oriented Disentanglement and Augmentation Modeling Scheme for Better Resisting Subpopulation ShiftsZiquan Ou, Zijun ZhangNeurIPS 2024 · 1 citation
- Proximal-Based Generative Modeling for Bayesian Inverse ProblemsBoyang Zhang, Zhiguo Wang, Ya-Feng LiuICML 2026
Builds on8
- NVAE: A Deep Hierarchical Variational AutoencoderArash Vahdat, Jan KautzNeurIPS 2020 · 1,141 citations
- From Variational to Deterministic AutoencodersPartha Ghosh, Mehdi S. M. Sajjadi, Antonio Vergari, Michael J. Black et al.ICLR 2020 · 298 citations
- Demystifying Inter-Class DisentanglementAviv Gabbay, Yedid HoshenICLR 2020 · 63 citations
- Learning Autoencoders with Relational RegularizationHongteng Xu, Dixin Luo, Ricardo Henao, Svati Shah et al.ICML 2020 · 47 citations
- A U-Net Based Discriminator for Generative Adversarial NetworksEdgar Schönfeld, Bernt Schiele, Anna KhorevaCVPR 2020
Related papers
- IntroVNMT: An Introspective Model for Variational Neural Machine TranslationXin Sheng, Linli Xu, Junliang Guo, Jingchang Liu et al.AAAI 2020 · 6 citations
- The Autoencoding Variational AutoencoderA. Taylan Cemgil, Sumedh Ghaisas, Krishnamurthy Dvijotham, Sven Gowal et al.NeurIPS 2020 · 81 citations
- Unsupervised Object Representation Learning using Translation and Rotation Group Equivariant VAEAlireza Nasiri, Tristan BeplerNeurIPS 2022 · 18 citations
- Explicitly Minimizing the Blur Error of Variational AutoencodersGustav Bredell, Kyriakos Flouris, Krishna Chaitanya, Ertunc Erdil et al.ICLR 2023 · 8 citations
- Dual Contradistinctive Generative AutoencoderGaurav Parmar, Dacheng Li, Kwonjoon Lee, Zhuowen TuCVPR 2021
