When Is Unsupervised Disentanglement Possible?
Daniella Horan, Eitan Richardson, Yair Weiss
Abstract
A common assumption in many domains is that high dimensional data are a smooth nonlinear function of a small number of independent factors. When is it possible to recover the factors from unlabeled data? In the context of deep models this problem is called "disentanglement" and was recently shown to be impossible without additional strong assumptions [17, 19] . In this paper, we show that the assumption of local isometry together with non-Gaussianity of the factors, is sufficient to provably recover disentangled representations from data. We leverage recent advances in deep generative models to construct manifolds of highly realistic images for which the ground truth latent representation is known, and test whether modern and classical methods succeed in recovering the latent factors. For many different manifolds, we find that a spectral method that explicitly optimizes local isometry and non-Gaussianity consistently finds the correct latent factors, while baseline deep autoencoders do not. We propose how to encourage deep autoencoders to find encodings that satisfy local isometry and show that this helps them discover disentangled representations. Overall, our results suggest that in some realistic settings, unsupervised disentanglement is provably possible, without any domain-specific assumptions. 𝑧! 𝑧! 𝑧! 𝑧! 𝑧" 𝑧" 𝑧" 𝑧" 𝑦 ! 𝑦 " 𝑦 ! 𝑦 " (a) cat1 (b) HLLE+ICA (c) IRMAE
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers22
- Function Classes for Identifiable Nonlinear Independent Component AnalysisSimon Buchholz, Michel Besserve, Bernhard SchölkopfNeurIPS 2022 · 63 citations
- Additive Decoders for Latent Variables Identification and Cartesian-Product ExtrapolationSébastien Lachapelle, Divyat Mahajan, Ioannis Mitliagkas, Simon Lacoste-JulienNeurIPS 2023 · 61 citations
- Disentanglement via Latent QuantizationKyle Hsu, William Dorrell, James C. R. Whittington, Jiajun Wu et al.NeurIPS 2023 · 54 citations
- Provable Compositional Generalization for Object-Centric LearningThaddäus Wiedemer, Jack Brady, Alexander Panfilov, Attila Juhos et al.ICLR 2024 · 40 citations
- From Causal to Concept-Based Representation LearningGoutham Rajendran, Simon Buchholz, Bryon Aragam, Bernhard Schölkopf et al.NeurIPS 2024 · 37 citations
Builds on4
- GANSpace: Discovering Interpretable GAN ControlsErik Härkönen, Aaron Hertzmann, Jaakko Lehtinen, Sylvain ParisNeurIPS 2020 · 1,049 citations
- Theory and Evaluation Metrics for Learning Disentangled RepresentationsKien Do, Truyen TranICLR 2020 · 107 citations
- Implicit Rank-Minimizing AutoencoderLi Jing, Jure Zbontar, Yann LeCunNeurIPS 2020 · 63 citations
- Analyzing and Improving the Image Quality of StyleGANTero Karras, Samuli Laine, Miika Aittala, Janne Hellsten et al.CVPR 2020
Related papers
- Learning disentangled representations via product manifold projectionMarco Fumero, Luca Cosmo, Simone Melzi, Emanuele RodolàICML 2021 · 29 citations
- Unsupervised Disentanglement Without Compromises : How Functional Orthogonality Enforces IdentifiabilityMathieu Simon, Pascal Frossard, Christophe De VleeschouwerICML 2026
- Multifactor Sequential Disentanglement via Structured Koopman AutoencodersNimrod Berman, Ilan Naiman, Omri AzencotICLR 2023 · 4 citations
- Weakly Supervised Disentanglement by Pairwise SimilaritiesJunxiang Chen, Kayhan BatmanghelichAAAI 2020 · 59 citations
- DisUnknown: Distilling Unknown Factors for Disentanglement LearningSitao Xiang, Yuming Gu, Pengda Xiang, Menglei Chai et al.ICCV 2021 · 6 citations
