Understanding the Generalization Benefit of Model Invariance from a Data Perspective
Sicheng Zhu, Bang An, Furong Huang
Abstract
Machine learning models that are developed with invariance to certain types of data transformations have demonstrated superior generalization performance in practice. However, the underlying mechanism that explains why invariance leads to better generalization is not well-understood, limiting our ability to select appropriate data transformations for a given dataset. This paper studies the generalization benefit of model invariance by introducing the sample cover induced by transformations, i.e., a representative subset of a dataset that can approximately recover the whole dataset using transformations. Based on this notion, we refine the generalization bound for invariant models and characterize the suitability of a set of data transformations by the sample covering number induced by transformations, i.e., the smallest size of its induced sample covers. We show that the generalization bound can be tightened for suitable transformations that have a small sample covering number. Moreover, our proposed sample covering number can be empirically evaluated, providing a practical guide for selecting transformations to develop model invariance for better generalization. We evaluate the sample covering numbers for commonly used transformations on multiple datasets and demonstrate that the smaller sample covering number for a set of transformations indicates a smaller gap between the test and training error for invariant models, thus validating our propositions.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 803518b6-a7d6-4339-924f-066badb9b3a9Cited by top-tier papers10
- PAC-Bayes Compression Bounds So Tight That They Can Explain GeneralizationSanae Lotfi, Marc Finzi, Sanyam Kapoor, Andres Potapczynski et al.NeurIPS 2022 · 98 citations
- Approximation-Generalization Trade-offs under (Approximate) Group EquivarianceMircea Petrache, Shubhendu TrivediNeurIPS 2023 · 54 citations
- Transferring Fairness under Distribution Shifts via Fair Consistency RegularizationBang An, Zora Che, Mucong Ding, Furong HuangNeurIPS 2022 · 44 citations
- On the Strong Correlation Between Model Invariance and GeneralizationWeijian Deng, Stephen Gould, Liang ZhengNeurIPS 2022 · 28 citations
- C-Disentanglement: Discovering Causally-Independent Generative Factors under an Inductive Bias of ConfounderXiaoyu Liu, Jiaxin Yuan, Bang An, Yuancheng Xu et al.NeurIPS 2023 · 13 citations
Builds on7
- Unsupervised Data Augmentation for Consistency TrainingQizhe Xie, Zihang Dai, Eduard H. Hovy, Thang Luong et al.NeurIPS 2020 · 2,774 citations
- Distributionally Robust Neural NetworksShiori Sagawa, Pang Wei Koh, Tatsunori B. Hashimoto, Percy LiangICLR 2020 · 1,578 citations
- A Group-Theoretic Framework for Data AugmentationShuxiao Chen, Edgar Dobriban, Jane H. LeeNeurIPS 2020 · 254 citations
- Model-Based Domain GeneralizationAlexander Robey, George J. Pappas, Hamed HassaniNeurIPS 2021 · 167 citations
- Provably Strict Generalisation Benefit for Equivariant ModelsBryn Elesedy, Sheheryar ZaidiICML 2021 · 100 citations
Related papers
- A PAC-Bayesian Generalization Bound for Equivariant NetworksArash Behboodi, Gabriele Cesa, Taco S. CohenNeurIPS 2022 · 26 citations
- Capacity of Group-invariant Linear Readouts from Equivariant Representations: How Many Objects can be Linearly Classified Under All Possible Views?Matthew Farrell, Blake Bordelon, Shubhendu Trivedi, Cengiz PehlevanICLR 2022 · 6 citations
- The Exact Sample Complexity Gain from Invariances for Kernel RegressionBehrooz Tahmasebi, Stefanie JegelkaNeurIPS 2023 · 29 citations
- Adversarial Robust Generalization of Graph Neural NetworksChang Cao, Han Li, Yulong Wang, Rui Wu et al.ICML 2025
- Symmetries in PAC-Bayesian LearningArmin Beck, Peter OchsICML 2026 · 1 citation
