Consistency Regularization for Variational Auto-Encoders
Samarth Sinha, Adji Bousso Dieng
Abstract
Variational auto-encoders ( s) are a powerful approach to unsupervised learning. They enable scalable approximate posterior inference in latent-variable models using variational inference ( ). A posits a variational family parameterized by a deep neural network-called an encoder-that takes data as input. This encoder is shared across all the observations, which amortizes the cost of inference. However the encoder of a has the undesirable property that it maps a given observation and a semantics-preserving transformation of it to different latent representations. This "inconsistency" of the encoder lowers the quality of the learned representations, especially for downstream tasks, and also negatively affects generalization. In this paper, we propose a regularization method to enforce consistency in s. The idea is to minimize the Kullback-Leibler ( ) divergence between the variational distribution when conditioning on the observation and the variational distribution when conditioning on a random semantic-preserving transformation of this observation. This regularization is applicable to any . In our experiments we apply it to four different variants on several benchmark datasets and found it always improves the quality of the learned representations but also leads to better generalization. In particular, when applied to the nouveau variational auto-encoder ( ), our regularization method yields state-of-the-art performance on , -10, and . We also applied our method to 3D data and found it learns representations of superior quality as measured by accuracy on a downstream classification task. Finally, we show our method can even outperform the triplet loss, an advanced and popular contrastive learning-based method for representation learning. 1
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5aba7b39-d55f-4a36-bac1-7594444d03d7Cited by top-tier papers21
- Maximum Likelihood Training of Score-Based Diffusion ModelsYang Song, Conor Durkan, Iain Murray, Stefano ErmonNeurIPS 2021 · 958 citations
- Can We Leave Deepfake Data Behind in Training Deepfake Detector?Jikang Cheng, Zhiyuan Yan, Ying Zhang, Yuhao Luo et al.NeurIPS 2024 · 85 citations
- Aligning Optimization Trajectories with Diffusion Models for Constrained Design GenerationGiorgio Giannone, Akash Srivastava, Ole Winther, Faez AhmedNeurIPS 2023 · 79 citations
- Densely connected normalizing flowsMatej Grcic, Ivan Grubisic, Sinisa SegvicNeurIPS 2021 · 67 citations
- Maximum Likelihood Training of Implicit Nonlinear Diffusion ModelDongjun Kim, Byeonghu Na, Se Jung Kwon, Dongsoo Lee et al.NeurIPS 2022 · 61 citations
Builds on5
- Unsupervised Data Augmentation for Consistency TrainingQizhe Xie, Zihang Dai, Eduard H. Hovy, Thang Luong et al.NeurIPS 2020 · 2,774 citations
- Image Augmentation Is All You Need: Regularizing Deep Reinforcement Learning from PixelsDenis Yarats, Ilya Kostrikov, Rob FergusICLR 2021 · 911 citations
- Variational Adversarial Active LearningSamarth Sinha, Sayna Ebrahimi, Trevor DarrellICCV 2019 · 662 citations
- Consistency Regularization for Generative Adversarial NetworksHan Zhang, Zizhao Zhang, Augustus Odena, Honglak LeeICLR 2020 · 305 citations
- Distribution Augmentation for Generative ModelingHeewoo Jun, Rewon Child, Mark Chen, John Schulman et al.ICML 2020 · 68 citations
Related papers
- Vector Quantization-Based Regularization for AutoencodersHanwei Wu, Markus FlierlAAAI 2020 · 33 citations
- The Autoencoding Variational AutoencoderA. Taylan Cemgil, Sumedh Ghaisas, Krishnamurthy Dvijotham, Sven Gowal et al.NeurIPS 2020 · 81 citations
- Adversarially Robust Representations with Smooth EncodersA. Taylan Cemgil, Sumedh Ghaisas, Krishnamurthy (Dj) Dvijotham, Pushmeet KohliICLR 2020 · 34 citations
- Learning Optimal Priors for Task-Invariant Representations in Variational AutoencodersHiroshi Takahashi, Tomoharu Iwata, Atsutoshi Kumagai, Sekitoshi Kanai et al.KDD 2022 · 4 citations
- Multimodal Gaussian Mixture Variational Autoencoder with Consistency RegularizationsYarui Chen, Lehan Hong, Jianlin Shao, Jianning Yang et al.AAAI 2026
