Lost in Latent Space: Examining failures of disentangled models at combinatorial generalisation
Milton Llera Montero, Jeffrey S. Bowers, Rui Ponte Costa, Casimir J. H. Ludwig, Gaurav Malhotra
摘要
Recent research has shown that generative models with highly disentangled representations fail to generalise to unseen combination of generative factor values. These findings contradict earlier research which showed improved performance in out-of-training distribution settings when compared to entangled representations. Additionally, it is not clear if the reported failures are due to (a) encoders failing to map novel combinations to the proper regions of the latent space, or (b) novel combinations being mapped correctly but the decoder being unable to render the correct output for the unseen combinations. We investigate these alternatives by testing several models on a range of datasets and training settings. We find that (i) when models fail, their encoders also fail to map unseen combinations to correct regions of the latent space and (ii) when models succeed, it is either because the test conditions do not exclude enough examples, or because excluded cases involve combinations of object properties with its shape. We argue that to generalise properly, models not only need to capture factors of variation, but also understand how to invert the process that causes the visual input.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Compositional Generalization from First PrinciplesThaddäus Wiedemer, Prasanna Mayilvahanan, Matthias Bethge, Wieland BrendelNeurIPS 2023 · 被引用 78 次
- Provable Compositional Generalization for Object-Centric LearningThaddäus Wiedemer, Jack Brady, Alexander Panfilov, Attila Juhos 等ICLR 2024 · 被引用 40 次
- Scaling can lead to compositional generalizationFlorian Redhardt, Yassir Akram, Simon SchugNeurIPS 2025 · 被引用 11 次
- Quantum latent distributions in deep generative modelsOmar Bacarreza, Thorin Farnsworth, Alexander Makarovskiy, Hugo Wallner 等ICML 2026 · 被引用 7 次
- Enriching Disentanglement: From Logical Definitions to Quantitative MetricsYivan Zhang, Masashi SugiyamaNeurIPS 2024 · 被引用 4 次
它引用的顶会 Paper7
- Weakly-Supervised Disentanglement Without CompromisesFrancesco Locatello, Ben Poole, Gunnar Rätsch, Bernhard Schölkopf 等ICML 2020 · 被引用 361 次
- Towards Nonlinear Disentanglement in Natural Data with Temporal Sparse CodingDavid A. Klindt, Lukas Schott, Yash Sharma, Ivan Ustyuzhaninov 等ICLR 2021 · 被引用 156 次
- InfoGAN-CR and ModelCentrality: Self-supervised Model Training and Selection for Disentangling GANsZinan Lin, Kiran Koshy Thekumparampil, Giulia Fanti, Sewoong OhICML 2020 · 被引用 106 次
- The role of Disentanglement in GeneralisationMilton Llera Montero, Casimir J. H. Ludwig, Rui Ponte Costa, Gaurav Malhotra 等ICLR 2021 · 被引用 97 次
- Unsupervised Model Selection for Variational Disentangled Representation LearningSunny Duan, Loic Matthey, Andre Saraiva, Nick Watters 等ICLR 2020 · 被引用 87 次
相关 Paper
- Compositional Generalization via Forced Rendering of Disentangled LatentsQiyao Liang, Daoyuan Qian, Liu Ziyin, Ila R. FieteICML 2025
- MAGANet: Achieving Combinatorial Generalization by Modeling a Group ActionGeonho Hwang, Jaewoong Choi, Hyunsoo Cho, Myungjoo KangICML 2023 · 被引用 4 次
- Unsupervised Causal Generative Understanding of ImagesTitas Anciukevicius, Patrick Fox-Roberts, Edward Rosten, Paul HendersonNeurIPS 2022 · 被引用 6 次
- On the Transfer of Disentangled Representations in Realistic SettingsAndrea Dittadi, Frederik Träuble, Francesco Locatello, Manuel Wuthrich 等ICLR 2021 · 被引用 17 次
- GIRAFFE: Representing Scenes As Compositional Generative Neural Feature FieldsMichael Niemeyer, Andreas GeigerCVPR 2021
