A Tale of Two Symmetries: Exploring the Loss Landscape of Equivariant Models
Yuqing Xie, Tess E. Smidt
Abstract
Equivariant neural networks have proven to be effective for tasks with known underlying symmetries. However, optimizing equivariant networks can be tricky and best training practices are less established than for standard networks. In particular, recent works have found small training benefits from relaxing equivariance constraints. This raises the question: do equivariance constraints introduce fundamental obstacles to optimization? Or do they simply require different hyperparameter tuning? In this work, we investigate this question through a theoretical analysis of the loss landscape geometry. We focus on networks built using permutation representations, which we can view as a subset of unconstrained MLPs. Importantly, we show that the parameter symmetries of the unconstrained model has nontrivial effects on the loss landscape of the equivariant subspace and under certain conditions can provably prevent learning of the global minima. Further, we empirically demonstrate in such cases, relaxing to an unconstrained MLP can sometimes solve the issue. Interestingly, the weights eventually found via relaxation corresponds to a different choice of group representation in the hidden layer. From this, we draw 3 key takeaways. (1) By viewing the unconstrained version of an architecture, we can uncover hidden parameter symmetries which were broken by choice of constraint enforcement (2) Hidden symmetries give important insights on loss landscapes and can induce critical points and even minima (3) Hidden symmetry induced minima can sometimes be escaped by constraint relaxation and we observe the network jumps to a different choice of constraint enforcement. Effective equivariance relaxation may require rethinking the fixed choice of group representation in the hidden layers.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers5
- Achieving Approximate Symmetry Is Exponentially Easier than Exact SymmetryBehrooz Tahmasebi, Melanie WeberICLR 2026 · 8 citations
- Revisiting the Canonicalization for Fast and Accurate Crystal Tensor Property PredictionHaowei Hua, Jingwen Yang, Wanyu Lin, Pan ZhouAAAI 2026 · 1 citation
- Identifiable Equivariant Networks are Layerwise EquivariantVahid Shahverdi, Giovanni Luca Marchetti, Georg Bökman, Kathlén KohnICML 2026
- Recurrent Equivariant Constraint Modulation: Learning Per-Layer Symmetry Relaxation from DataStefanos Pertigkiozoglou, Mircea Petrache, Shubhendu Trivedi, Kostas DaniilidisICML 2026
- Globscope: Toward a Global View of the Loss LandscapeMashiat Mustaq, Xavier M.CVPR 2026
Builds on23
- MACE: Higher Order Equivariant Message Passing Neural Networks for Fast and Accurate Force FieldsIlyes Batatia, Dávid Péter Kovács, Gregor N. C. Simm, Christoph Ortner et al.NeurIPS 2022 · 1,448 citations
- Equivariant Diffusion for Molecule Generation in 3DEmiel Hoogeboom, Victor Garcia Satorras, Clément Vignac, Max WellingICML 2022 · 865 citations
- Linear Mode Connectivity and the Lottery Ticket HypothesisJonathan Frankle, Gintare Karolina Dziugaite, Daniel M. Roy, Michael CarbinICML 2020 · 750 citations
- The Role of Permutation Invariance in Linear Mode Connectivity of Neural NetworksRahim Entezari, Hanie Sedghi, Olga Saukh, Behnam NeyshaburICLR 2022 · 301 citations
- A Practical Method for Constructing Equivariant Multilayer Perceptrons for Arbitrary Matrix GroupsMarc Finzi, Max Welling, Andrew Gordon WilsonICML 2021 · 226 citations
Related papers
- Improving Equivariant Model Training via Constraint RelaxationStefanos Pertigkiozoglou, Evangelos Chatzipantazis, Shubhendu Trivedi, Kostas DaniilidisNeurIPS 2024 · 26 citations
- Relaxing Equivariance Constraints with Non-stationary Continuous FiltersTycho F. A. van der Ouderaa, David W. Romero, Mark van der WilkNeurIPS 2022 · 51 citations
- Equivariance-aware Architectural Optimization of Neural NetworksKaitlin Maile, Dennis George Wilson, Patrick ForréICLR 2023
- Learning (Approximately) Equivariant Networks via Constrained OptimizationAndrei Manolache, Luiz F. O. Chamon, Mathias NiepertNeurIPS 2025 · 12 citations
- Universal Equivariant Multilayer PerceptronsSiamak RavanbakhshICML 2020 · 60 citations
