Identifiable Equivariant Networks are Layerwise Equivariant
Vahid Shahverdi, Giovanni Luca Marchetti, Georg Bökman, Kathlén Kohn
Abstract
We investigate the relation between end-to-end equivariance and layerwise equivariance in deep neural networks. We prove the following: For a network whose end-to-end function is equivariant with respect to group actions on the input and output spaces, there is a parameter choice yielding the same end-to-end function such that its layers are equivariant with respect to some group actions on the latent spaces. Our result assumes that the parameters of the model are identifiable in an appropriate sense. This identifiability property has been established in the literature for a large class of networks, to which our results apply immediately, while it is conjectural for others. The theory we develop is grounded in an abstract formalism, and is therefore architecture-agnostic. Overall, our results provide a mathematical explanation for the emergence of equivariant structures in the weights of neural networks during training -- a phenomenon that is consistently observed in practice.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3eb75812-2ddf-47dd-9019-d9f5452d905fBuilds on15
- On the Symmetries of Deep Learning Models and their Internal RepresentationsCharles Godfrey, Davis Brown, Tegan Emerson, Henry KvingeNeurIPS 2022 · 78 citations
- Universal Equivariant Multilayer PerceptronsSiamak RavanbakhshICML 2020 · 60 citations
- Functional vs. parametric equivalence of ReLU networksMary Phuong, Christoph H. LampertICLR 2020 · 53 citations
- Pure and Spurious Critical Points: a Geometric Study of Linear NetworksMatthew Trager, Kathlén Kohn, Joan BrunaICLR 2020 · 41 citations
- Hidden Symmetries of ReLU NetworksJ. Elisenda Grigsby, Kathryn Lindsey, David RolnickICML 2023 · 35 citations
Related papers
- Investigating how ReLU-networks encode symmetriesGeorg Bökman, Fredrik KahlNeurIPS 2023 · 11 citations
- Emergent Equivariance in Deep EnsemblesJan E. Gerken, Pan KesselICML 2024 · 14 citations
- Unsupervised Learning of Group Invariant and Equivariant RepresentationsRobin Winter, Marco Bertolini, Tuan Le, Frank Noé et al.NeurIPS 2022 · 61 citations
- Learning Layer-wise Equivariances Automatically using GradientsTycho F. A. van der Ouderaa, Alexander Immer, Mark van der WilkNeurIPS 2023 · 28 citations
- How Jellyfish Characterise Alternating Group Equivariant Neural NetworksEdward Pearce-CrumpICML 2023 · 5 citations
