A Tale of Two Symmetries: Exploring the Loss Landscape of Equivariant Models
Yuqing Xie, Tess E. Smidt
摘要
Equivariant neural networks have proven to be effective for tasks with known underlying symmetries. However, optimizing equivariant networks can be tricky and best training practices are less established than for standard networks. In particular, recent works have found small training benefits from relaxing equivariance constraints. This raises the question: do equivariance constraints introduce fundamental obstacles to optimization? Or do they simply require different hyperparameter tuning? In this work, we investigate this question through a theoretical analysis of the loss landscape geometry. We focus on networks built using permutation representations, which we can view as a subset of unconstrained MLPs. Importantly, we show that the parameter symmetries of the unconstrained model has nontrivial effects on the loss landscape of the equivariant subspace and under certain conditions can provably prevent learning of the global minima. Further, we empirically demonstrate in such cases, relaxing to an unconstrained MLP can sometimes solve the issue. Interestingly, the weights eventually found via relaxation corresponds to a different choice of group representation in the hidden layer. From this, we draw 3 key takeaways. (1) By viewing the unconstrained version of an architecture, we can uncover hidden parameter symmetries which were broken by choice of constraint enforcement (2) Hidden symmetries give important insights on loss landscapes and can induce critical points and even minima (3) Hidden symmetry induced minima can sometimes be escaped by constraint relaxation and we observe the network jumps to a different choice of constraint enforcement. Effective equivariance relaxation may require rethinking the fixed choice of group representation in the hidden layers.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Achieving Approximate Symmetry Is Exponentially Easier than Exact SymmetryBehrooz Tahmasebi, Melanie WeberICLR 2026 · 被引用 8 次
- Revisiting the Canonicalization for Fast and Accurate Crystal Tensor Property PredictionHaowei Hua, Jingwen Yang, Wanyu Lin, Pan ZhouAAAI 2026 · 被引用 1 次
- Identifiable Equivariant Networks are Layerwise EquivariantVahid Shahverdi, Giovanni Luca Marchetti, Georg Bökman, Kathlén KohnICML 2026
- Recurrent Equivariant Constraint Modulation: Learning Per-Layer Symmetry Relaxation from DataStefanos Pertigkiozoglou, Mircea Petrache, Shubhendu Trivedi, Kostas DaniilidisICML 2026
- Globscope: Toward a Global View of the Loss LandscapeMashiat Mustaq, Xavier M.CVPR 2026
它引用的顶会 Paper23
- MACE: Higher Order Equivariant Message Passing Neural Networks for Fast and Accurate Force FieldsIlyes Batatia, Dávid Péter Kovács, Gregor N. C. Simm, Christoph Ortner 等NeurIPS 2022 · 被引用 1,448 次
- Equivariant Diffusion for Molecule Generation in 3DEmiel Hoogeboom, Victor Garcia Satorras, Clément Vignac, Max WellingICML 2022 · 被引用 865 次
- Linear Mode Connectivity and the Lottery Ticket HypothesisJonathan Frankle, Gintare Karolina Dziugaite, Daniel M. Roy, Michael CarbinICML 2020 · 被引用 750 次
- The Role of Permutation Invariance in Linear Mode Connectivity of Neural NetworksRahim Entezari, Hanie Sedghi, Olga Saukh, Behnam NeyshaburICLR 2022 · 被引用 301 次
- A Practical Method for Constructing Equivariant Multilayer Perceptrons for Arbitrary Matrix GroupsMarc Finzi, Max Welling, Andrew Gordon WilsonICML 2021 · 被引用 226 次
相关 Paper
- Improving Equivariant Model Training via Constraint RelaxationStefanos Pertigkiozoglou, Evangelos Chatzipantazis, Shubhendu Trivedi, Kostas DaniilidisNeurIPS 2024 · 被引用 26 次
- Relaxing Equivariance Constraints with Non-stationary Continuous FiltersTycho F. A. van der Ouderaa, David W. Romero, Mark van der WilkNeurIPS 2022 · 被引用 51 次
- Equivariance-aware Architectural Optimization of Neural NetworksKaitlin Maile, Dennis George Wilson, Patrick ForréICLR 2023
- Learning (Approximately) Equivariant Networks via Constrained OptimizationAndrei Manolache, Luiz F. O. Chamon, Mathias NiepertNeurIPS 2025 · 被引用 12 次
- Universal Equivariant Multilayer PerceptronsSiamak RavanbakhshICML 2020 · 被引用 60 次
