Approximation-Generalization Trade-offs under (Approximate) Group Equivariance
Mircea Petrache, Shubhendu Trivedi
摘要
The explicit incorporation of task-specific inductive biases through symmetry has emerged as a general design precept in the development of high-performance machine learning models. For example, group equivariant neural networks have demonstrated impressive performance across various domains and applications such as protein and drug design. A prevalent intuition about such models is that the integration of relevant symmetry results in enhanced generalization. Moreover, it is posited that when the data and/or the model may only exhibit or symmetry, the optimal or best-performing model is one where the model symmetry aligns with the data symmetry. In this paper, we conduct a formal unified investigation of these intuitions. To begin, we present general quantitative bounds that demonstrate how models capturing task-specific symmetries lead to improved generalization. In fact, our results do not require the transformations to be finite or even form a group and can work with partial or approximate equivariance. Utilizing this quantification, we examine the more general question of model mis-specification i.e. when the model symmetries don't align with the data symmetries. We establish, for a given symmetry group, a quantitative comparison between the approximate/partial equivariance of the model and that of the data distribution, precisely connecting model equivariance error and data equivariance error. Our result delineates conditions under which the model equivariance error is optimal, thereby yielding the best-performing model for the given task and data. Our results are the most general results of their type in the literature.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper25
- Approximately Equivariant Graph NetworksNingyuan Huang, Ron Levie, Soledad VillarNeurIPS 2023 · 被引用 29 次
- Improving Equivariant Model Training via Constraint RelaxationStefanos Pertigkiozoglou, Evangelos Chatzipantazis, Shubhendu Trivedi, Kostas DaniilidisNeurIPS 2024 · 被引用 26 次
- Discovering Symmetry Breaking in Physical Systems with Relaxed Group ConvolutionRui Wang, Elyssa F. Hofgard, Hang Gao, Robin Walters 等ICML 2024 · 被引用 20 次
- Probing Equivariance and Symmetry Breaking in Convolutional NetworksSharvaree Vadgama, Mohammad Mohaiminul Islam, Domas Buracas, Christian Shewmake 等NeurIPS 2025 · 被引用 15 次
- On the hardness of learning under symmetriesBobak T. Kiani, Thien Le, Hannah Lawrence, Stefanie Jegelka 等ICLR 2024 · 被引用 14 次
它引用的顶会 Paper18
- Directional Message Passing for Molecular GraphsJohannes Klicpera, Janek Groß, Stephan GünnemannICLR 2020 · 被引用 1,079 次
- A Group-Theoretic Framework for Data AugmentationShuxiao Chen, Edgar Dobriban, Jane H. LeeNeurIPS 2020 · 被引用 254 次
- A Practical Method for Constructing Equivariant Multilayer Perceptrons for Arbitrary Matrix GroupsMarc Finzi, Max Welling, Andrew Gordon WilsonICML 2021 · 被引用 226 次
- B-Spline CNNs on Lie groupsErik J. BekkersICLR 2020 · 被引用 155 次
- Permutation Equivariant Models for Compositional Generalization in LanguageJonathan Gordon, David Lopez-Paz, Marco Baroni, Diane BouchacourtICLR 2020 · 被引用 112 次
相关 Paper
- MatrixNet: Learning over symmetry groups using learned group representationsLucas Laird, Circe Hsu, Asilata Bapat, Robin WaltersNeurIPS 2024 · 被引用 2 次
- The Surprising Effectiveness of Equivariant Models in Domains with Latent SymmetryDian Wang, Jung Yeon Park, Neel Sortur, Lawson L. S. Wong 等ICLR 2023 · 被引用 2 次
- Learning Partial Equivariances From DataDavid W. Romero, Suhas LohitNeurIPS 2022 · 被引用 54 次
- A PAC-Bayesian Generalization Bound for Equivariant NetworksArash Behboodi, Gabriele Cesa, Taco S. CohenNeurIPS 2022 · 被引用 26 次
- Any-Subgroup Equivariant Networks via Symmetry BreakingAbhinav Goel, Derek Lim, Hannah Lawrence, Stefanie Jegelka 等ICLR 2026
