Any-Subgroup Equivariant Networks via Symmetry Breaking
Abhinav Goel, Derek Lim, Hannah Lawrence, Stefanie Jegelka, Ningyuan Huang
摘要
The inclusion of symmetries as an inductive bias, known as equivariance, often improves generalization on geometric data (e.g. grids, sets, and graphs). However, equivariant architectures are usually highly constrained, designed for symmetries chosen a priori, and not applicable to datasets with other symmetries. This precludes the development of flexible, multi-modal foundation models capable of processing diverse data equivariantly. In this work, we build a single model --- the Any-Subgroup Equivariant Network (ASEN) --- that can be simultaneously equivariant to several groups, simply by modulating a certain auxiliary input feature. In particular, we start with a fully permutation-equivariant base model, and then obtain subgroup equivariance by using a symmetry-breaking input whose automorphism group is that subgroup. However, finding an input with the desired automorphism group is computationally hard. We overcome this by relaxing from exact to approximate symmetry breaking, leveraging the notion of 2-closure to derive fast algorithms. Theoretically, we show that our subgroup-equivariant networks can simulate equivariant MLPs, and their universality can be guaranteed if the base model is universal. Empirically, we validate our method on symmetry selection for graph and image tasks, as well as multitask and transfer learning for sequence tasks, showing that a single network equivariant to multiple permutation subgroups outperforms both separate equivariant models and a single non-equivariant model.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper22
- Long Range Arena : A Benchmark for Efficient TransformersYi Tay, Mostafa Dehghani, Samira Abnar, Yikang Shen 等ICLR 2021 · 被引用 881 次
- What graph neural networks cannot learn: depth vs widthAndreas LoukasICLR 2020 · 被引用 336 次
- A Practical Method for Constructing Equivariant Multilayer Perceptrons for Arbitrary Matrix GroupsMarc Finzi, Max Welling, Andrew Gordon WilsonICML 2021 · 被引用 226 次
- Approximately Equivariant Networks for Imperfectly Symmetric DynamicsRui Wang, Robin Walters, Rose YuICML 2022 · 被引用 111 次
- Provably Strict Generalisation Benefit for Equivariant ModelsBryn Elesedy, Sheheryar ZaidiICML 2021 · 被引用 100 次
相关 Paper
- Learning Probabilistic Symmetrization for Architecture Agnostic EquivarianceJinwoo Kim, Dat Nguyen, Ayhan Suleymanzade, Hyeokjun An 等NeurIPS 2023 · 被引用 32 次
- Universal Equivariant Multilayer PerceptronsSiamak RavanbakhshICML 2020 · 被引用 60 次
- Meta-learning Symmetries by ReparameterizationAllan Zhou, Tom Knowles, Chelsea FinnICLR 2021 · 被引用 105 次
- Learning Symmetric Embeddings for Equivariant World ModelsJung Yeon Park, Ondrej Biza, Linfeng Zhao, Jan-Willem van de Meent 等ICML 2022 · 被引用 56 次
- Approximation-Generalization Trade-offs under (Approximate) Group EquivarianceMircea Petrache, Shubhendu TrivediNeurIPS 2023 · 被引用 54 次
