Consistency Regularization for Variational Auto-Encoders
Samarth Sinha, Adji Bousso Dieng
摘要
Variational auto-encoders ( s) are a powerful approach to unsupervised learning. They enable scalable approximate posterior inference in latent-variable models using variational inference ( ). A posits a variational family parameterized by a deep neural network-called an encoder-that takes data as input. This encoder is shared across all the observations, which amortizes the cost of inference. However the encoder of a has the undesirable property that it maps a given observation and a semantics-preserving transformation of it to different latent representations. This "inconsistency" of the encoder lowers the quality of the learned representations, especially for downstream tasks, and also negatively affects generalization. In this paper, we propose a regularization method to enforce consistency in s. The idea is to minimize the Kullback-Leibler ( ) divergence between the variational distribution when conditioning on the observation and the variational distribution when conditioning on a random semantic-preserving transformation of this observation. This regularization is applicable to any . In our experiments we apply it to four different variants on several benchmark datasets and found it always improves the quality of the learned representations but also leads to better generalization. In particular, when applied to the nouveau variational auto-encoder ( ), our regularization method yields state-of-the-art performance on , -10, and . We also applied our method to 3D data and found it learns representations of superior quality as measured by accuracy on a downstream classification task. Finally, we show our method can even outperform the triplet loss, an advanced and popular contrastive learning-based method for representation learning. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper21
- Maximum Likelihood Training of Score-Based Diffusion ModelsYang Song, Conor Durkan, Iain Murray, Stefano ErmonNeurIPS 2021 · 被引用 958 次
- Can We Leave Deepfake Data Behind in Training Deepfake Detector?Jikang Cheng, Zhiyuan Yan, Ying Zhang, Yuhao Luo 等NeurIPS 2024 · 被引用 85 次
- Aligning Optimization Trajectories with Diffusion Models for Constrained Design GenerationGiorgio Giannone, Akash Srivastava, Ole Winther, Faez AhmedNeurIPS 2023 · 被引用 79 次
- Densely connected normalizing flowsMatej Grcic, Ivan Grubisic, Sinisa SegvicNeurIPS 2021 · 被引用 67 次
- Maximum Likelihood Training of Implicit Nonlinear Diffusion ModelDongjun Kim, Byeonghu Na, Se Jung Kwon, Dongsoo Lee 等NeurIPS 2022 · 被引用 61 次
它引用的顶会 Paper5
- Unsupervised Data Augmentation for Consistency TrainingQizhe Xie, Zihang Dai, Eduard H. Hovy, Thang Luong 等NeurIPS 2020 · 被引用 2,774 次
- Image Augmentation Is All You Need: Regularizing Deep Reinforcement Learning from PixelsDenis Yarats, Ilya Kostrikov, Rob FergusICLR 2021 · 被引用 911 次
- Variational Adversarial Active LearningSamarth Sinha, Sayna Ebrahimi, Trevor DarrellICCV 2019 · 被引用 662 次
- Consistency Regularization for Generative Adversarial NetworksHan Zhang, Zizhao Zhang, Augustus Odena, Honglak LeeICLR 2020 · 被引用 305 次
- Distribution Augmentation for Generative ModelingHeewoo Jun, Rewon Child, Mark Chen, John Schulman 等ICML 2020 · 被引用 68 次
相关 Paper
- Vector Quantization-Based Regularization for AutoencodersHanwei Wu, Markus FlierlAAAI 2020 · 被引用 33 次
- The Autoencoding Variational AutoencoderA. Taylan Cemgil, Sumedh Ghaisas, Krishnamurthy Dvijotham, Sven Gowal 等NeurIPS 2020 · 被引用 81 次
- Adversarially Robust Representations with Smooth EncodersA. Taylan Cemgil, Sumedh Ghaisas, Krishnamurthy (Dj) Dvijotham, Pushmeet KohliICLR 2020 · 被引用 34 次
- Learning Optimal Priors for Task-Invariant Representations in Variational AutoencodersHiroshi Takahashi, Tomoharu Iwata, Atsutoshi Kumagai, Sekitoshi Kanai 等KDD 2022 · 被引用 4 次
- Multimodal Gaussian Mixture Variational Autoencoder with Consistency RegularizationsYarui Chen, Lehan Hong, Jianlin Shao, Jianning Yang 等AAAI 2026
