ID and OOD Performance Are Sometimes Inversely Correlated on Real-world Datasets
Damien Teney, Yong Lin, Seong Joon Oh, Ehsan Abbasnejad
摘要
Context. Several studies have compared the in-distribution (ID) and out-ofdistribution (OOD) performance of models in computer vision and NLP. They report a frequent positive correlation and some surprisingly never even observe an inverse correlation indicative of a necessary trade-off. The possibility of inverse patterns is important to determine whether ID performance can serve as a proxy for OOD generalization capabilities. Findings. This paper shows with multiple datasets that inverse correlations between ID and OOD performance do happen in real-world data -not only in theoretical worst-case settings. We also explain theoretically how these cases can arise even in a minimal linear setting, and why past studies could miss such cases due to a biased selection of models. Implications. Our observations lead to recommendations that contradict those found in much of the current literature. • High OOD performance sometimes requires trading off ID performance. • Focusing on ID performance alone may not lead to optimal OOD performance. It may produce diminishing (eventually negative) returns in OOD performance. • In these cases, studies on OOD generalization that use ID performance for model selection (a common recommended practice) will necessarily miss the bestperforming models, making these studies blind to a whole range of phenomena.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- Understanding and Improving Feature Learning for Out-of-Distribution GeneralizationYongqiang Chen, Wei Huang, Kaiwen Zhou, Yatao Bian 等NeurIPS 2023 · 被引用 49 次
- Evading the Simplicity Bias: Training a Diverse Set of Models Discovers Solutions with Superior OOD GeneralizationDamien Teney, Ehsan Abbasnejad, Simon Lucey, Anton van den HengelCVPR 2022 · 被引用 32 次
- The Entropy Enigma: Success and Failure of Entropy MinimizationOri Press, Ravid Shwartz-Ziv, Yann LeCun, Matthias BethgeICML 2024 · 被引用 27 次
- Understanding the detrimental class-level effects of data augmentationPolina Kirichenko, Mark Ibrahim, Randall Balestriero, Diane Bouchacourt 等NeurIPS 2023 · 被引用 25 次
- Effective Robustness against Natural Distribution Shifts for Models with Different Training DataZhouxing Shi, Nicholas Carlini, Ananth Balashankar, Ludwig Schmidt 等NeurIPS 2023 · 被引用 17 次
它引用的顶会 Paper28
- WILDS: A Benchmark of in-the-Wild Distribution ShiftsPang Wei Koh, Shiori Sagawa, Henrik Marklund, Sang Michael Xie 等ICML 2021 · 被引用 1,773 次
- In Search of Lost Domain GeneralizationIshaan Gulrajani, David Lopez-PazICLR 2021 · 被引用 1,416 次
- Measuring Robustness to Natural Distribution Shifts in Image ClassificationRohan Taori, Achal Dave, Vaishaal Shankar, Nicholas Carlini 等NeurIPS 2020 · 被引用 731 次
- The Pitfalls of Simplicity Bias in Neural NetworksHarshay Shah, Kaustav Tamuly, Aditi Raghunathan, Prateek Jain 等NeurIPS 2020 · 被引用 503 次
- The Risks of Invariant Risk MinimizationElan Rosenfeld, Pradeep Kumar Ravikumar, Andrej RisteskiICLR 2021 · 被引用 356 次
相关 Paper
- Assaying Out-Of-Distribution Generalization in Transfer LearningFlorian Wenzel, Andrea Dittadi, Peter V. Gehler, Carl-Johann Simon-Gabriel 等NeurIPS 2022 · 被引用 93 次
- Aggregation Hides Out-of-Distribution Generalization Failures from Spurious CorrelationsOlawale Salaudeen, Haoran Zhang, Kumail Alhamoud, Sara Beery 等NeurIPS 2025 · 被引用 3 次
- Accuracy on the Curve: On the Nonlinear Correlation of ML Performance Between Data SubpopulationsWeixin Liang, Yining Mao, Yongchan Kwon, Xinyu Yang 等ICML 2023 · 被引用 7 次
- On the Adversarial Robustness of Out-of-distribution Generalization ModelsXin Zou, Weiwei LiuNeurIPS 2023 · 被引用 10 次
- Two Sides of Meta-Learning Evaluation: In vs. Out of DistributionAmrith Setlur, Oscar Li, Virginia SmithNeurIPS 2021 · 被引用 17 次
