Demystifying Inductive Biases for (Beta-)VAE Based Architectures
Dominik Zietlow, Michal Rolínek, Georg Martius
摘要
The performance of -Variational-Autoencoders (-VAEs) and their variants on learning semantically meaningful, disentangled representations is unparalleled. On the other hand, there are theoretical arguments suggesting the impossibility of unsupervised disentanglement. In this work, we shed light on the inductive bias responsible for the success of VAE-based architectures. We show that in classical datasets the structure of variance, induced by the generating factors, is conveniently aligned with the latent directions fostered by the VAE objective. This builds the pivotal bias on which the disentangling abilities of VAEs rely. By small, elaborate perturbations of existing datasets, we hide the convenient correlation structure that is easily exploited by a variety of architectures. To demonstrate this, we construct modified versions of standard datasets in which (i) the generative factors are perfectly preserved; (ii) each image undergoes a mild transformation causing a small change of variance; (iii) the leading VAE-based disentanglement architectures fail to produce disentangled representations whilst the performance of a non-variational method remains unchanged. The construction of our modifications is nontrivial and relies on recent progress on mechanistic understanding of -VAEs and their connection to PCA. We strengthen that connection by providing additional insights that are of stand-alone interest.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Function Classes for Identifiable Nonlinear Independent Component AnalysisSimon Buchholz, Michel Besserve, Bernhard SchölkopfNeurIPS 2022 · 被引用 63 次
- Disentanglement with Biological Constraints: A Theory of Functional Cell TypesJames C. R. Whittington, Will Dorrell, Surya Ganguli, Timothy BehrensICLR 2023 · 被引用 13 次
- Robustness of Nonlinear Representation LearningSimon Buchholz, Bernhard SchölkopfICML 2024 · 被引用 11 次
- A Survey of Inductive Reasoning for Large Language ModelsKedi Chen, Dezhao Ruan, Yuhao Dan, Yaoting Wang 等ACL 2026 · 被引用 5 次
- Geometric Inductive Biases for Identifiable Unsupervised Learning of Disentangled RepresentationsZiqi Pan, Li Niu, Liqing ZhangAAAI 2023 · 被引用 3 次
它引用的顶会 Paper4
- Towards Nonlinear Disentanglement in Natural Data with Temporal Sparse CodingDavid A. Klindt, Lukas Schott, Yash Sharma, Ivan Ustyuzhaninov 等ICLR 2021 · 被引用 156 次
- Unsupervised Model Selection for Variational Disentangled Representation LearningSunny Duan, Loic Matthey, Andre Saraiva, Nick Watters 等ICLR 2020 · 被引用 87 次
- A Theory of Independent Mechanisms for Extrapolation in Generative ModelsMichel Besserve, Rémy Sun, Dominik Janzing, Bernhard SchölkopfAAAI 2021 · 被引用 27 次
- Towards Unsupervised Learning of Generative Models for 3D Controllable Image SynthesisYiyi Liao, Katja Schwarz, Lars M. Mescheder, Andreas GeigerCVPR 2020
相关 Paper
- Why do Variational Autoencoders Really Promote Disentanglement?Pratik Bhowal, Achint Soni, Sirisha RambhatlaICML 2024 · 被引用 11 次
- Improving VAEs' Robustness to Adversarial AttackMatthew Willetts, Alexander Camuto, Tom Rainforth, Stephen J. Roberts 等ICLR 2021 · 被引用 30 次
- Local Disentanglement in Variational Auto-Encoders Using Jacobian RegularizationTravers Rhodes, Daniel D. LeeNeurIPS 2021 · 被引用 24 次
- Adversarial Disentanglement with Grouped ObservationsJózsef NémethAAAI 2020 · 被引用 8 次
- The role of Disentanglement in GeneralisationMilton Llera Montero, Casimir J. H. Ludwig, Rui Ponte Costa, Gaurav Malhotra 等ICLR 2021 · 被引用 97 次
