Soft-IntroVAE: Analyzing and Improving the Introspective Variational Autoencoder
Tal Daniel, Aviv Tamar
摘要
The recently introduced introspective variational autoencoder (IntroVAE) exhibits outstanding image generations, and allows for amortized inference using an image encoder. The main idea in IntroVAE is to train a VAE adversarially, using the VAE encoder to discriminate between generated and real data samples. However, the original In-troVAE loss function relied on a particular hinge-loss formulation that is very hard to stabilize in practice, and its theoretical convergence analysis ignored important terms in the loss. In this work, we take a step towards better understanding of the IntroVAE model, its practical implementation, and its applications. We propose the Soft-IntroVAE, a modified IntroVAE that replaces the hinge-loss terms with a smooth exponential loss on generated samples. This change significantly improves training stability, and also enables theoretical analysis of the complete algorithm. Interestingly, we show that the IntroVAE converges to a distribution that minimizes a sum of KL distance from the data distribution and an entropy term. We discuss the implications of this result, and demonstrate that it induces competitive image generation and reconstruction. Finally, we describe an application of Soft-IntroVAE to unsupervised image translation, and demonstrate compelling results. Code and additional information is available on the project websitetaldatech.github.io/soft-intro-vae-web.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Unsupervised Image Representation Learning with Deep Latent ParticlesTal Daniel, Aviv TamarICML 2022 · 被引用 18 次
- Property Existence Inference against Generative ModelsLijin Wang, Jingjing Wang, Jie Wan, Lin Long 等USENIX Security 2024 · 被引用 13 次
- CODA: A Correlation-Oriented Disentanglement and Augmentation Modeling Scheme for Better Resisting Subpopulation ShiftsZiquan Ou, Zijun ZhangNeurIPS 2024 · 被引用 1 次
- Proximal-Based Generative Modeling for Bayesian Inverse ProblemsBoyang Zhang, Zhiguo Wang, Ya-Feng LiuICML 2026
它引用的顶会 Paper8
- NVAE: A Deep Hierarchical Variational AutoencoderArash Vahdat, Jan KautzNeurIPS 2020 · 被引用 1,141 次
- From Variational to Deterministic AutoencodersPartha Ghosh, Mehdi S. M. Sajjadi, Antonio Vergari, Michael J. Black 等ICLR 2020 · 被引用 298 次
- Demystifying Inter-Class DisentanglementAviv Gabbay, Yedid HoshenICLR 2020 · 被引用 63 次
- Learning Autoencoders with Relational RegularizationHongteng Xu, Dixin Luo, Ricardo Henao, Svati Shah 等ICML 2020 · 被引用 47 次
- A U-Net Based Discriminator for Generative Adversarial NetworksEdgar Schönfeld, Bernt Schiele, Anna KhorevaCVPR 2020
相关 Paper
- IntroVNMT: An Introspective Model for Variational Neural Machine TranslationXin Sheng, Linli Xu, Junliang Guo, Jingchang Liu 等AAAI 2020 · 被引用 6 次
- The Autoencoding Variational AutoencoderA. Taylan Cemgil, Sumedh Ghaisas, Krishnamurthy Dvijotham, Sven Gowal 等NeurIPS 2020 · 被引用 81 次
- Unsupervised Object Representation Learning using Translation and Rotation Group Equivariant VAEAlireza Nasiri, Tristan BeplerNeurIPS 2022 · 被引用 18 次
- Explicitly Minimizing the Blur Error of Variational AutoencodersGustav Bredell, Kyriakos Flouris, Krishna Chaitanya, Ertunc Erdil 等ICLR 2023 · 被引用 8 次
- Dual Contradistinctive Generative AutoencoderGaurav Parmar, Dacheng Li, Kwonjoon Lee, Zhuowen TuCVPR 2021
