Neural Networks for Learning Counterfactual G-Invariances from Single Environments
S. Chandra Mouli, Bruno Ribeiro
摘要
Despite -or maybe because of-their astonishing capacity to fit data, neural networks are believed to have difficulties extrapolating beyond training data distribution. This work shows that, for extrapolations based on finite transformation groups, a model's inability to extrapolate is unrelated to its capacity. Rather, the shortcoming is inherited from a learning hypothesis: Examples not explicitly observed with infinitely many training examples have underspecified outcomes in the learner's model. In order to endow neural networks with the ability to extrapolate over group transformations, we introduce a learning framework counterfactually-guided by the learning hypothesis that any group invariance to (known) transformation groups is mandatory even without evidence, unless the learner deems it inconsistent with the training data. Unlike existing invariance-driven methods for (counterfactual) extrapolations, this framework allows extrapolations from a single environment. Finally, we introduce sequence and image extrapolation tasks that validate our framework and showcase the shortcomings of traditional approaches. INTRODUCTION Neural networks are widely praised for their ability to interpolate the training data. However, in some applications, they have also been shown to be unable to learn patterns that can provably extrapolate out-of-distribution (beyond the training data distribution) (
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Size-Invariant Graph Representations for Graph Classification ExtrapolationsBeatrice Bevilacqua, Yangze Zhou, Bruno RibeiroICML 2021 · 被引用 124 次
- Learning Probabilistic Symmetrization for Architecture Agnostic EquivarianceJinwoo Kim, Dat Nguyen, Ayhan Suleymanzade, Hyeokjun An 等NeurIPS 2023 · 被引用 32 次
- Causal Effect Regularization: Automated Detection and Removal of Spurious CorrelationsAbhinav Kumar, Amit Deshpande, Amit SharmaNeurIPS 2023 · 被引用 7 次
- Group Downsampling with Equivariant Anti-aliasingMd Ashiqur Rahman, Raymond A. YehICLR 2025
它引用的顶会 Paper6
- Out-of-Distribution Generalization via Risk Extrapolation (REx)David Krueger, Ethan Caballero, Jörn-Henrik Jacobsen, Amy Zhang 等ICML 2021 · 被引用 1,163 次
- A Meta-Transfer Objective for Learning to Disentangle Causal MechanismsYoshua Bengio, Tristan Deleu, Nasim Rahaman, Nan Rosemary Ke 等ICLR 2020 · 被引用 371 次
- How Neural Networks Extrapolate: From Feedforward to Graph Neural NetworksKeyulu Xu, Mozhi Zhang, Jingling Li, Simon Shaolei Du 等ICLR 2021 · 被引用 364 次
- The Risks of Invariant Risk MinimizationElan Rosenfeld, Pradeep Kumar Ravikumar, Andrej RisteskiICLR 2021 · 被引用 356 次
- A Group-Theoretic Framework for Data AugmentationShuxiao Chen, Edgar Dobriban, Jane H. LeeNeurIPS 2020 · 被引用 254 次
相关 Paper
- First Steps Toward Understanding the Extrapolation of Nonlinear Models to Unseen DomainsKefan Dong, Tengyu MaICLR 2023 · 被引用 3 次
- Learning Representations that Support ExtrapolationTaylor W. Webb, Zachary Dulberg, Steven Frankland, Alexander A. Petrov 等ICML 2020 · 被引用 60 次
- Explainable Neural Rule LearningShaoyun Shi, Yuexiang Xie, Zhen Wang, Bolin Ding 等WWW 2022 · 被引用 12 次
- Location Attention for Extrapolation to Longer SequencesYann Dubois, Gautier Dagan, Dieuwke Hupkes, Elia BruniACL 2020 · 被引用 2 次
- Neural Networks and the Chomsky HierarchyGrégoire Delétang, Anian Ruoss, Jordi Grau-Moya, Tim Genewein 等ICLR 2023 · 被引用 45 次
