Interaction Asymmetry: A General Principle for Learning Composable Abstractions
Jack Brady, Julius von Kügelgen, Sébastien Lachapelle, Simon Buchholz, Thomas Kipf, Wieland Brendel
摘要
Learning disentangled representations of concepts and re-composing them in unseen ways is crucial for generalizing to out-of-domain situations. However, the underlying properties of concepts that enable such disentanglement and compositional generalization remain poorly understood. In this work, we propose the principle of interaction asymmetry which states: "Parts of the same concept have more complex interactions than parts of different concepts". We formalize this via block diagonality conditions on the th order derivatives of the generator mapping concepts to observed data, where different orders of "complexity" correspond to different . Using this formalism, we prove that interaction asymmetry enables both disentanglement and compositional generalization. Our results unify recent theoretical results for learning concepts of objects, which we show are recovered as special cases with or . We provide results for up to , thus extending these prior works to more flexible generator functions, and conjecture that the same proof strategies generalize to larger . Practically, our theory suggests that, to disentangle concepts, an autoencoder should penalize its latent capacity and the interactions between concepts during decoding. We propose an implementation of these criteria using a flexible Transformer-based VAE, with a novel regularizer on the attention weights of the decoder. On synthetic image datasets consisting of objects, we provide evidence that this model can achieve comparable object disentanglement to existing models that use more explicit object-centric priors.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- Scaling can lead to compositional generalizationFlorian Redhardt, Yassir Akram, Simon SchugNeurIPS 2025 · 被引用 11 次
- Diverse Influence Component Analysis: A Geometric Approach to Nonlinear Mixture IdentifiabilityHoang-Son Nguyen, Xiao FuNeurIPS 2025 · 被引用 6 次
- From Isolation to Entanglement: When Do Interpretability Methods Identify and Disentangle Known Concepts?Aaron Mueller, Andrew Lee, Shruti Joshi, Ekdeep Singh Lubana 等ACL 2026 · 被引用 5 次
- Mechanistic Independence: A Principle for Identifiable Disentangled RepresentationsStefan Matthes, Zhiwei Han, Hao ShenICLR 2026 · 被引用 3 次
- Scalable Evaluation and Neural Models for Compositional GeneralizationGiacomo Camposampiero, Pietro Barbiero, Michael Hersche, Roger Wattenhofer 等NeurIPS 2025 · 被引用 3 次
它引用的顶会 Paper51
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao 等ICCV 2023 · 被引用 13,211 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
- Zero-Shot Text-to-Image GenerationAditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray 等ICML 2021 · 被引用 6,356 次
- Object-Centric Learning with Slot AttentionFrancesco Locatello, Dirk Weissenborn, Thomas Unterthiner, Aravindh Mahendran 等NeurIPS 2020 · 被引用 1,275 次
- Perceiver IO: A General Architecture for Structured Inputs & OutputsAndrew Jaegle, Sebastian Borgeaud, Jean-Baptiste Alayrac, Carl Doersch 等ICLR 2022 · 被引用 797 次
相关 Paper
- Provable Compositional Generalization for Object-Centric LearningThaddäus Wiedemer, Jack Brady, Alexander Panfilov, Attila Juhos 等ICLR 2024 · 被引用 40 次
- The role of Disentanglement in GeneralisationMilton Llera Montero, Casimir J. H. Ludwig, Rui Ponte Costa, Gaurav Malhotra 等ICLR 2021 · 被引用 97 次
- Efficient Iterative Amortized Inference for Learning Symmetric and Disentangled Multi-Object RepresentationsPatrick Emami, Pan He, Sanjay Ranka, Anand RangarajanICML 2021 · 被引用 48 次
- Compositional Generalization in Unsupervised Compositional Representation Learning: A Study on Disentanglement and Emergent LanguageZhenlin Xu, Marc Niethammer, Colin RaffelNeurIPS 2022 · 被引用 59 次
- Unsupervised Disentanglement Without Compromises : How Functional Orthogonality Enforces IdentifiabilityMathieu Simon, Pascal Frossard, Christophe De VleeschouwerICML 2026
