Distributional Autoencoders Know the Score
Andrej Leban
Abstract
The Distributional Principal Autoencoder (DPA) combines distributionally correct reconstruction with principal-component-like interpretability of the encodings. In this work, we provide exact theoretical guarantees on both fronts. First, we derive a closed-form relation linking each optimal level-set geometry to the data-distribution score. This result explains DPA's empirical ability to disentangle factors of variation of the data, as well as allows the score to be recovered directly from samples. When the data follows the Boltzmann distribution, we demonstrate that this relation yields an approximation of the minimum free-energy path for the Mueller-Brown potential in a single fit. Second, we prove that if the data lies on a manifold that can be approximated by the encoder, latent components beyond the manifold dimension are conditionally independent of the data distribution - carrying no additional information - and thus reveal the intrinsic dimension. Together, these results show that a single model can learn the data distribution and its intrinsic dimension with exact guarantees simultaneously, unifying two longstanding goals of unsupervised learning.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ad8cbdca-9a30-4a4c-8457-b8777edefebeBuilds on2
- Score-Based Generative Modeling through Stochastic Differential EquationsYang Song, Jascha Sohl-Dickstein, Diederik P. Kingma, Abhishek Kumar et al.ICLR 2021 · 1,270 citations
- Learning Manifold Dimensions with Conditional Variational AutoencodersYijia Zheng, Tong He, Yixuan Qiu, David P. WipfNeurIPS 2022 · 34 citations
Related papers
- Score-based Pullback Riemannian Geometry: Extracting the Data Manifold Geometry using Anisotropic FlowsWillem Diepeveen, Georgios Batzolis, Zakhar Shumaylov, Carola-Bibiane SchönliebICML 2025
- Geometric Inductive Biases for Identifiable Unsupervised Learning of Disentangled RepresentationsZiqi Pan, Li Niu, Liqing ZhangAAAI 2023 · 3 citations
- Marginalization is not Marginal: No Bad VAE Local Minima when Learning Optimal Sparse RepresentationsDavid WipfICML 2023 · 5 citations
- Disentanglement Learning via TopologyNikita Balabin, Daria Voronkova, Ilya Trofimov, Evgeny Burnaev et al.ICML 2024 · 4 citations
- Quantitative Understanding of VAE as a Non-linearly Scaled Isometric EmbeddingAkira Nakagawa, Keizo Kato, Taiji SuzukiICML 2021 · 10 citations
