Provably Strict Generalisation Benefit for Equivariant Models
Bryn Elesedy, Sheheryar Zaidi
Abstract
It is widely believed that engineering a model to be invariant/equivariant improves generalisation. Despite the growing popularity of this approach, a precise characterisation of the generalisation benefit is lacking. By considering the simplest case of linear models, this paper provides the first provably non-zero improvement in generalisation for invariant/equivariant models when the target distribution is invariant/equivariant with respect to a compact group. Moreover, our work reveals an interesting relationship between generalisation, the number of training examples and properties of the group action. Our results rest on an observation of the structure of function spaces under averaging operators which, along with its consequences for feature averaging, may be of independent interest.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 7ecc3ac3-5166-4e46-a368-e917130a5c36Cited by top-tier papers45
- SE(3) diffusion model with application to protein backbone generationJason Yim, Brian L. Trippe, Valentin De Bortoli, Emile Mathieu et al.ICML 2023 · 313 citations
- Scalars are universal: Equivariant machine learning, structured like classical physicsSoledad Villar, David W. Hogg, Kate Storey-Fisher, Weichi Yao et al.NeurIPS 2021 · 185 citations
- Equivariant Architectures for Learning in Deep Weight SpacesAviv Navon, Aviv Shamsian, Idan Achituve, Ethan Fetaya et al.ICML 2023 · 101 citations
- PAC-Bayes Compression Bounds So Tight That They Can Explain GeneralizationSanae Lotfi, Marc Finzi, Sanyam Kapoor, Andres Potapczynski et al.NeurIPS 2022 · 98 citations
- Residual Pathway Priors for Soft Equivariance ConstraintsMarc Finzi, Greg Benton, Andrew Gordon WilsonNeurIPS 2021 · 89 citations
Related papers
- Provably Strict Generalisation Benefit for Invariance in Kernel MethodsBryn ElesedyNeurIPS 2021 · 35 citations
- Generalization Bounds for Canonicalization: A Comparative Study with Group AveragingBehrooz Tahmasebi, Stefanie JegelkaICLR 2025
- The Exact Sample Complexity Gain from Invariances for Kernel RegressionBehrooz Tahmasebi, Stefanie JegelkaNeurIPS 2023 · 29 citations
- A PAC-Bayesian Generalization Bound for Equivariant NetworksArash Behboodi, Gabriele Cesa, Taco S. CohenNeurIPS 2022 · 26 citations
- Symmetries in PAC-Bayesian LearningArmin Beck, Peter OchsICML 2026 · 1 citation
