A Statistical Analysis of Wasserstein Autoencoders for Intrinsically Low-dimensional Data
Saptarshi Chakraborty, Peter L. Bartlett
Abstract
Variational Autoencoders (VAEs) have gained significant popularity among researchers as a powerful tool for understanding unknown distributions based on limited samples. This popularity stems partly from their impressive performance and partly from their ability to provide meaningful feature representations in the latent space. Wasserstein Autoencoders (WAEs), a variant of VAEs, aim to not only improve model efficiency but also interpretability. However, there has been limited focus on analyzing their statistical guarantees. The matter is further complicated by the fact that the data distributions to which WAEs are applied - such as natural images - are often presumed to possess an underlying low-dimensional structure within a high-dimensional feature space, which current theory does not adequately account for, rendering known bounds inefficient. To bridge the gap between the theory and practice of WAEs, in this paper, we show that WAEs can learn the data distributions when the network architectures are properly chosen. We show that the convergence rates of the expected excess risk in the number of samples for WAEs are independent of the high feature dimension, instead relying only on the intrinsic dimension of the data distribution.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 879080dc-1c9b-40a2-80aa-25950263f1b3Builds on6
- Zero-Shot Text-to-Image GenerationAditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray et al.ICML 2021 · 6,356 citations
- The Intrinsic Dimension of Images and Its Impact on LearningPhillip Pope, Chen Zhu, Ahmed Abdelkader, Micah Goldblum et al.ICLR 2021 · 381 citations
- On Deep Generative Models for Approximation and Estimation of Distributions on ManifoldsBiraj Dahal, Alexander Havrilla, Minshuo Chen, Tuo Zhao et al.NeurIPS 2022 · 17 citations
- Variational autoencoders in the presence of low-dimensional data: landscape and implicit biasFrederic Koehler, Viraj Mehta, Chenghui Zhou, Andrej RisteskiICLR 2022 · 14 citations
- Statistical Regeneration Guarantees of the Wasserstein Autoencoder with Latent Space ConsistencyAnish Chakrabarty, Swagatam DasNeurIPS 2021 · 10 citations
Related papers
- Information-Theoretic Generalization Bounds for VAEs: A Role of Encoder and Latent VariableFutoshi Futami, Masahiro FujisawaICML 2026
- StrWAEs to Invariant RepresentationsHyunjong Lee, Yedarm Seong, Sungdong Lee, Joong-Ho WonICML 2024
- Statistical Guarantees for Variational Autoencoders using PAC-Bayesian TheorySokhna Diarra Mbacke, Florence Clerc, Pascal GermainNeurIPS 2023 · 22 citations
- Learning Manifold Dimensions with Conditional Variational AutoencodersYijia Zheng, Tong He, Yixuan Qiu, David P. WipfNeurIPS 2022 · 34 citations
- Shape your Space: A Gaussian Mixture Regularization Approach to Deterministic AutoencodersAmrutha Saseendran, Kathrin Skubch, Stefan Falkner, Margret KeuperNeurIPS 2021 · 13 citations
