Identifying Metric Structures of Deep Latent Variable Models
Stas Syrota, Yevgen Zainchkovskyy, Johnny Xi, Benjamin Bloem-Reddy, Søren Hauberg
Abstract
Deep latent variable models learn condensed representations of data that, hopefully, reflect the inner workings of the studied phenomena. Unfortunately, these latent representations are not statistically identifiable, meaning they cannot be uniquely determined. Domain experts, therefore, need to tread carefully when interpreting these. Current solutions limit the lack of identifiability through additional constraints on the latent variable model, e.g. by requiring labeled training data, or by restricting the expressivity of the model. We change the goal: instead of identifying the latent variables, we identify relationships between them such as meaningful distances, angles, and volumes. We prove this is feasible under very mild model conditions and without additional labeled data. We empirically demonstrate that our theory results in more reliable latent distances, offering a principled path forward in extracting trustworthy conclusions from deep latent variable models.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c66074b1-269d-4e34-84ab-de8058f3b8e5Cited by top-tier papers2
- Connecting Neural Models Latent Geometries with Relative Geodesic RepresentationsHanlin Yu, Berfin Inal, Georgios Arvanitidis, Søren Hauberg et al.NeurIPS 2025 · 5 citations
- Contact Wasserstein Geodesics for Non-Conservative Schrödinger BridgesAndrea Testa, Søren Hauberg, Tamim Asfour, Leonel RozoICLR 2026
Builds on11
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Weakly-Supervised Disentanglement Without CompromisesFrancesco Locatello, Ben Poole, Gunnar Rätsch, Bernhard Schölkopf et al.ICML 2020 · 361 citations
- Flows for simultaneous manifold learning and density estimationJohann Brehmer, Kyle CranmerNeurIPS 2020 · 187 citations
- Weakly Supervised Disentanglement with GuaranteesRui Shu, Yining Chen, Abhishek Kumar, Stefano Ermon et al.ICLR 2020 · 148 citations
- Identifiability of deep generative models without auxiliary informationBohdan Kivva, Goutham Rajendran, Pradeep Ravikumar, Bryon AragamNeurIPS 2022 · 87 citations
Related papers
- A Critical Look at the Consistency of Causal Estimation with Deep Latent Variable ModelsSeveri Rissanen, Pekka MarttinenNeurIPS 2021 · 38 citations
- Unsupervised Disentanglement Without Compromises : How Functional Orthogonality Enforces IdentifiabilityMathieu Simon, Pascal Frossard, Christophe De VleeschouwerICML 2026
- Posterior Collapse and Latent Variable Non-identifiabilityYixin Wang, David M. Blei, John P. CunninghamNeurIPS 2021 · 97 citations
- Identifying through Flows for Recovering Latent RepresentationsShen Li, Bryan Hooi, Gim Hee LeeICLR 2020 · 15 citations
- Self-Supervised Learning with Data Augmentations Provably Isolates Content from StyleJulius von Kügelgen, Yash Sharma, Luigi Gresele, Wieland Brendel et al.NeurIPS 2021 · 421 citations
