Neural Networks for Learning Counterfactual G-Invariances from Single Environments
S. Chandra Mouli, Bruno Ribeiro
Abstract
Despite -or maybe because of-their astonishing capacity to fit data, neural networks are believed to have difficulties extrapolating beyond training data distribution. This work shows that, for extrapolations based on finite transformation groups, a model's inability to extrapolate is unrelated to its capacity. Rather, the shortcoming is inherited from a learning hypothesis: Examples not explicitly observed with infinitely many training examples have underspecified outcomes in the learner's model. In order to endow neural networks with the ability to extrapolate over group transformations, we introduce a learning framework counterfactually-guided by the learning hypothesis that any group invariance to (known) transformation groups is mandatory even without evidence, unless the learner deems it inconsistent with the training data. Unlike existing invariance-driven methods for (counterfactual) extrapolations, this framework allows extrapolations from a single environment. Finally, we introduce sequence and image extrapolation tasks that validate our framework and showcase the shortcomings of traditional approaches. INTRODUCTION Neural networks are widely praised for their ability to interpolate the training data. However, in some applications, they have also been shown to be unable to learn patterns that can provably extrapolate out-of-distribution (beyond the training data distribution) (
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers4
- Size-Invariant Graph Representations for Graph Classification ExtrapolationsBeatrice Bevilacqua, Yangze Zhou, Bruno RibeiroICML 2021 · 124 citations
- Learning Probabilistic Symmetrization for Architecture Agnostic EquivarianceJinwoo Kim, Dat Nguyen, Ayhan Suleymanzade, Hyeokjun An et al.NeurIPS 2023 · 32 citations
- Causal Effect Regularization: Automated Detection and Removal of Spurious CorrelationsAbhinav Kumar, Amit Deshpande, Amit SharmaNeurIPS 2023 · 7 citations
- Group Downsampling with Equivariant Anti-aliasingMd Ashiqur Rahman, Raymond A. YehICLR 2025
Builds on6
- Out-of-Distribution Generalization via Risk Extrapolation (REx)David Krueger, Ethan Caballero, Jörn-Henrik Jacobsen, Amy Zhang et al.ICML 2021 · 1,163 citations
- A Meta-Transfer Objective for Learning to Disentangle Causal MechanismsYoshua Bengio, Tristan Deleu, Nasim Rahaman, Nan Rosemary Ke et al.ICLR 2020 · 371 citations
- How Neural Networks Extrapolate: From Feedforward to Graph Neural NetworksKeyulu Xu, Mozhi Zhang, Jingling Li, Simon Shaolei Du et al.ICLR 2021 · 364 citations
- The Risks of Invariant Risk MinimizationElan Rosenfeld, Pradeep Kumar Ravikumar, Andrej RisteskiICLR 2021 · 356 citations
- A Group-Theoretic Framework for Data AugmentationShuxiao Chen, Edgar Dobriban, Jane H. LeeNeurIPS 2020 · 254 citations
Related papers
- First Steps Toward Understanding the Extrapolation of Nonlinear Models to Unseen DomainsKefan Dong, Tengyu MaICLR 2023 · 3 citations
- Learning Representations that Support ExtrapolationTaylor W. Webb, Zachary Dulberg, Steven Frankland, Alexander A. Petrov et al.ICML 2020 · 60 citations
- Explainable Neural Rule LearningShaoyun Shi, Yuexiang Xie, Zhen Wang, Bolin Ding et al.WWW 2022 · 12 citations
- Location Attention for Extrapolation to Longer SequencesYann Dubois, Gautier Dagan, Dieuwke Hupkes, Elia BruniACL 2020 · 2 citations
- Neural Networks and the Chomsky HierarchyGrégoire Delétang, Anian Ruoss, Jordi Grau-Moya, Tim Genewein et al.ICLR 2023 · 45 citations
