ID and OOD Performance Are Sometimes Inversely Correlated on Real-world Datasets
Damien Teney, Yong Lin, Seong Joon Oh, Ehsan Abbasnejad
Abstract
Context. Several studies have compared the in-distribution (ID) and out-ofdistribution (OOD) performance of models in computer vision and NLP. They report a frequent positive correlation and some surprisingly never even observe an inverse correlation indicative of a necessary trade-off. The possibility of inverse patterns is important to determine whether ID performance can serve as a proxy for OOD generalization capabilities. Findings. This paper shows with multiple datasets that inverse correlations between ID and OOD performance do happen in real-world data -not only in theoretical worst-case settings. We also explain theoretically how these cases can arise even in a minimal linear setting, and why past studies could miss such cases due to a biased selection of models. Implications. Our observations lead to recommendations that contradict those found in much of the current literature. • High OOD performance sometimes requires trading off ID performance. • Focusing on ID performance alone may not lead to optimal OOD performance. It may produce diminishing (eventually negative) returns in OOD performance. • In these cases, studies on OOD generalization that use ID performance for model selection (a common recommended practice) will necessarily miss the bestperforming models, making these studies blind to a whole range of phenomena.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a64138ac-0ff1-4e06-bca1-5b2e72631b34Cited by top-tier papers12
- Understanding and Improving Feature Learning for Out-of-Distribution GeneralizationYongqiang Chen, Wei Huang, Kaiwen Zhou, Yatao Bian et al.NeurIPS 2023 · 49 citations
- Evading the Simplicity Bias: Training a Diverse Set of Models Discovers Solutions with Superior OOD GeneralizationDamien Teney, Ehsan Abbasnejad, Simon Lucey, Anton van den HengelCVPR 2022 · 32 citations
- The Entropy Enigma: Success and Failure of Entropy MinimizationOri Press, Ravid Shwartz-Ziv, Yann LeCun, Matthias BethgeICML 2024 · 27 citations
- Understanding the detrimental class-level effects of data augmentationPolina Kirichenko, Mark Ibrahim, Randall Balestriero, Diane Bouchacourt et al.NeurIPS 2023 · 25 citations
- Effective Robustness against Natural Distribution Shifts for Models with Different Training DataZhouxing Shi, Nicholas Carlini, Ananth Balashankar, Ludwig Schmidt et al.NeurIPS 2023 · 17 citations
Builds on28
- WILDS: A Benchmark of in-the-Wild Distribution ShiftsPang Wei Koh, Shiori Sagawa, Henrik Marklund, Sang Michael Xie et al.ICML 2021 · 1,773 citations
- In Search of Lost Domain GeneralizationIshaan Gulrajani, David Lopez-PazICLR 2021 · 1,416 citations
- Measuring Robustness to Natural Distribution Shifts in Image ClassificationRohan Taori, Achal Dave, Vaishaal Shankar, Nicholas Carlini et al.NeurIPS 2020 · 731 citations
- The Pitfalls of Simplicity Bias in Neural NetworksHarshay Shah, Kaustav Tamuly, Aditi Raghunathan, Prateek Jain et al.NeurIPS 2020 · 503 citations
- The Risks of Invariant Risk MinimizationElan Rosenfeld, Pradeep Kumar Ravikumar, Andrej RisteskiICLR 2021 · 356 citations
Related papers
- Assaying Out-Of-Distribution Generalization in Transfer LearningFlorian Wenzel, Andrea Dittadi, Peter V. Gehler, Carl-Johann Simon-Gabriel et al.NeurIPS 2022 · 93 citations
- Aggregation Hides Out-of-Distribution Generalization Failures from Spurious CorrelationsOlawale Salaudeen, Haoran Zhang, Kumail Alhamoud, Sara Beery et al.NeurIPS 2025 · 3 citations
- Accuracy on the Curve: On the Nonlinear Correlation of ML Performance Between Data SubpopulationsWeixin Liang, Yining Mao, Yongchan Kwon, Xinyu Yang et al.ICML 2023 · 7 citations
- On the Adversarial Robustness of Out-of-distribution Generalization ModelsXin Zou, Weiwei LiuNeurIPS 2023 · 10 citations
- Two Sides of Meta-Learning Evaluation: In vs. Out of DistributionAmrith Setlur, Oscar Li, Virginia SmithNeurIPS 2021 · 17 citations
