Investigating how ReLU-networks encode symmetries
Georg Bökman, Fredrik Kahl
Abstract
Many data symmetries can be described in terms of group equivariance and the most common way of encoding group equivariances in neural networks is by building linear layers that are group equivariant. In this work we investigate whether equivariance of a network implies that all layers are equivariant. On the theoretical side we find cases where equivariance implies layerwise equivariance, but also demonstrate that this is not the case generally. Nevertheless, we conjecture that CNNs that are trained to be equivariant will exhibit layerwise equivariance and explain how this conjecture is a weaker version of the recent permutation conjecture by Entezari et al. [2022] . We perform quantitative experiments with VGG-nets on CIFAR10 and qualitative experiments with ResNets on ImageNet to illustrate and support our theoretical findings. These experiments are not only of interest for understanding how group equivariance is encoded in ReLU-networks, but they also give a new perspective on Entezari et al.'s permutation conjecture as we find that it is typically easier to merge a network with a group-transformed version of itself than merging two different networks. .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext fdf20af5-e61e-48c8-a5cc-89bf30aac8dfCited by top-tier papers5
- The Empirical Impact of Neural Parameter Symmetries, or Lack ThereofDerek Lim, Theo (Moe) Putterman, Robin Walters, Haggai Maron et al.NeurIPS 2024 · 25 citations
- Symmetry Induces Structure and Constraint of LearningLiu ZiyinICML 2024 · 24 citations
- Identifiable Equivariant Networks are Layerwise EquivariantVahid Shahverdi, Giovanni Luca Marchetti, Georg Bökman, Kathlén KohnICML 2026
- Steerers: A Framework for Rotation Equivariant Keypoint DescriptorsGeorg Bökman, Johan Edstedt, Michael Felsberg, Fredrik KahlCVPR 2024
- Flopping for FLOPs: Leveraging Equivariance for Computational EfficiencyGeorg Bökman, David Nordström, Fredrik KahlICML 2025
Builds on17
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec et al.NeurIPS 2020 · 9,171 citations
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou et al.ICCV 2021 · 8,921 citations
- An Empirical Study of Training Self-Supervised Vision TransformersXinlei Chen, Saining Xie, Kaiming HeICCV 2021 · 2,340 citations
- SE(3)-Transformers: 3D Roto-Translation Equivariant Attention NetworksFabian Fuchs, Daniel E. Worrall, Volker Fischer, Max WellingNeurIPS 2020 · 1,025 citations
- Linear Mode Connectivity and the Lottery Ticket HypothesisJonathan Frankle, Gintare Karolina Dziugaite, Daniel M. Roy, Michael CarbinICML 2020 · 750 citations
Related papers
- A Characterization Theorem for Equivariant Networks with Point-wise ActivationsMarco Pacini, Xiaowen Dong, Bruno Lepri, Gabriele SantinICLR 2024 · 5 citations
- Universal Equivariant Multilayer PerceptronsSiamak RavanbakhshICML 2020 · 60 citations
- Implicit Bias of Linear Equivariant NetworksHannah Lawrence, Bobak Toussi Kiani, Kristian G. Georgiev, Andrew K. DienesICML 2022 · 18 citations
- Meta-learning Symmetries by ReparameterizationAllan Zhou, Tom Knowles, Chelsea FinnICLR 2021 · 105 citations
- Learning Partial Equivariances From DataDavid W. Romero, Suhas LohitNeurIPS 2022 · 54 citations
