Towards a Theoretical Framework of Out-of-Distribution Generalization
Haotian Ye, Chuanlong Xie, Tianle Cai, Ruichen Li, Zhenguo Li, Liwei Wang
摘要
Generalization to out-of-distribution (OOD) data is one of the central problems in modern machine learning. Recently, there is a surge of attempts to propose algorithms that mainly build upon the idea of extracting invariant features. Although intuitively reasonable, theoretical understanding of what kind of invariance can guarantee OOD generalization is still limited, and generalization to arbitrary out-of-distribution is clearly impossible. In this work, we take the first step towards rigorous and quantitative definitions of 1) what is OOD; and 2) what does it mean by saying an OOD problem is learnable. We also introduce a new concept of expansion function, which characterizes to what extent the variance is amplified in the test domains over the training domains, and therefore give a quantitative meaning of invariant features. Based on these, we prove OOD generalization error bounds. It turns out that OOD generalization largely depends on the expansion function. As recently pointed out by Gulrajani and Lopez-Paz (2020), any OOD learning algorithm without a model selection module is incomplete. Our theory naturally induces a model selection criterion. Extensive experiments on benchmark OOD datasets demonstrate that our model selection criterion has a significant advantage over baselines.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper45
- Learning Causal Semantic Representation for Out-of-Distribution PredictionChang Liu, Xinwei Sun, Jindong Wang, Haoyue Tang 等NeurIPS 2021 · 被引用 136 次
- MADG: Margin-based Adversarial Learning for Domain GeneralizationAveen Dayal, Vimal K. B., Linga Reddy Cenkeramaddi, C. Krishna Mohan 等NeurIPS 2023 · 被引用 102 次
- RankFeat: Rank-1 Feature Removal for Out-of-distribution DetectionYue Song, Nicu Sebe, Wei WangNeurIPS 2022 · 被引用 76 次
- Quantifying and Improving Transferability in Domain GeneralizationGuojun Zhang, Han Zhao, Yaoliang Yu, Pascal PoupartNeurIPS 2021 · 被引用 56 次
- A Theory of Label Propagation for Subpopulation ShiftTianle Cai, Ruiqi Gao, Jason D. Lee, Qi LeiICML 2021 · 被引用 55 次
它引用的顶会 Paper9
- Distributionally Robust Neural NetworksShiori Sagawa, Pang Wei Koh, Tatsunori B. Hashimoto, Percy LiangICLR 2020 · 被引用 1,578 次
- Out-of-Distribution Generalization via Risk Extrapolation (REx)David Krueger, Ethan Caballero, Jörn-Henrik Jacobsen, Amy Zhang 等ICML 2021 · 被引用 1,163 次
- Measuring Robustness to Natural Distribution Shifts in Image ClassificationRohan Taori, Achal Dave, Vaishaal Shankar, Nicholas Carlini 等NeurIPS 2020 · 被引用 731 次
- A Meta-Transfer Objective for Learning to Disentangle Causal MechanismsYoshua Bengio, Tristan Deleu, Nasim Rahaman, Nan Rosemary Ke 等ICLR 2020 · 被引用 371 次
- The Risks of Invariant Risk MinimizationElan Rosenfeld, Pradeep Kumar Ravikumar, Andrej RisteskiICLR 2021 · 被引用 356 次
相关 Paper
- Out-of-distribution Generalization with Causal Invariant TransformationsRuoyu Wang, Mingyang Yi, Zhitang Chen, Shengyu ZhuCVPR 2022 · 被引用 40 次
- HYPO: Hyperspherical Out-Of-Distribution GeneralizationHaoyue Bai, Yifei Ming, Julian Katz-Samuels, Yixuan LiICLR 2024 · 被引用 13 次
- Out-Of-Distribution Detection with Diversification (Provably)Haiyun Yao, Zongbo Han, Huazhu Fu, Xi Peng 等NeurIPS 2024 · 被引用 9 次
- Is Out-of-Distribution Detection Learnable?Zhen Fang, Yixuan Li, Jie Lu, Jiahua Dong 等NeurIPS 2022 · 被引用 188 次
- A Variational Information Theoretic Approach to Out-of-Distribution DetectionSudeepta Mondal, Zhuolin Jiang, Ganesh SundaramoorthiICML 2025
