Learning Manifold Dimensions with Conditional Variational Autoencoders
Yijia Zheng, Tong He, Yixuan Qiu, David P. Wipf
Abstract
Although the variational autoencoder (VAE) and its conditional extension (CVAE) are capable of state-of-the-art results across multiple domains, their precise behavior is still not fully understood, particularly in the context of data (like images) that lie on or near a low-dimensional manifold. For example, while prior work has suggested that the globally optimal VAE solution can learn the correct manifold dimension, a necessary (but not sufficient) condition for producing samples from the true data distribution, this has never been rigorously proven. Moreover, it remains unclear how such considerations would change when various types of conditioning variables are introduced, or when the data support is extended to a union of manifolds (e.g., as is likely the case for MNIST digits and related). In this work, we address these points by first proving that VAE global minima are indeed capable of recovering the correct manifold dimension. We then extend this result to more general CVAEs, demonstrating practical scenarios whereby the conditioning variables allow the model to adaptively learn manifolds of varying dimension across samples. Our analyses, which have practical implications for various CVAE design choices, are also supported by numerical results on both synthetic and real-world datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b153e54d-8e51-4275-9278-481f20acdc4dCited by top-tier papers8
- A Geometric View of Data Complexity: Efficient Local Intrinsic Dimension Estimation with Diffusion ModelsHamidreza Kamkari, Brendan Leigh Ross, Rasa Hosseinzadeh, Jesse C. Cresswell et al.NeurIPS 2024 · 49 citations
- Towards Understanding Future: Consistency Guided Probabilistic Modeling for Action AnticipationZhao Xie, Yadong Shi, Kewei Wu, Yaru Cheng et al.AAAI 2024 · 9 citations
- Marginalization is not Marginal: No Bad VAE Local Minima when Learning Optimal Sparse RepresentationsDavid WipfICML 2023 · 5 citations
- Distributional Autoencoders Know the ScoreAndrej LebanNeurIPS 2025 · 3 citations
- A Geometric Framework for Understanding Memorization in Generative ModelsBrendan Leigh Ross, Hamidreza Kamkari, Tongzi Wu, Rasa Hosseinzadeh et al.ICLR 2025
Builds on6
- The Intrinsic Dimension of Images and Its Impact on LearningPhillip Pope, Chen Zhu, Ahmed Abdelkader, Micah Goldblum et al.ICLR 2021 · 381 citations
- Probabilistic Transformer For Time Series AnalysisBinh Tang, David S. MattesonNeurIPS 2021 · 150 citations
- GRIN: Generative Relation and Intention Network for Multi-agent Trajectory PredictionLongyuan Li, Jian Yao, Li K. Wenliang, Tong He et al.NeurIPS 2021 · 51 citations
- On the Value of Infinite Gradients in Variational Autoencoder ModelsBin Dai, Wenliang Li, David P. WipfNeurIPS 2021 · 15 citations
- Marginalization is not Marginal: No Bad VAE Local Minima when Learning Optimal Sparse RepresentationsDavid WipfICML 2023 · 5 citations
Related papers
- Sparse Autoencoders, Again?Yin Lu, Xuening Zhu, Tong He, David WipfICML 2025
- Variational autoencoders in the presence of low-dimensional data: landscape and implicit biasFrederic Koehler, Viraj Mehta, Chenghui Zhou, Andrej RisteskiICLR 2022 · 14 citations
- Beyond Vanilla Variational Autoencoders: Detecting Posterior Collapse in Conditional and Hierarchical Variational AutoencodersHien Dang, Tho Tran Huu, Tan Minh Nguyen, Nhat HoICLR 2024 · 8 citations
- A Statistical Analysis of Wasserstein Autoencoders for Intrinsically Low-dimensional DataSaptarshi Chakraborty, Peter L. BartlettICLR 2024 · 3 citations
- Learning Optimal Priors for Task-Invariant Representations in Variational AutoencodersHiroshi Takahashi, Tomoharu Iwata, Atsutoshi Kumagai, Sekitoshi Kanai et al.KDD 2022 · 4 citations
