On the Importance of Feature Separability in Predicting Out-Of-Distribution Error
Renchunzi Xie, Hongxin Wei, Lei Feng, Yuzhou Cao, Bo An
Abstract
Estimating the generalization performance is practically challenging on out-of-distribution (OOD) data without ground-truth labels. While previous methods emphasize the connection between distribution difference and OOD accuracy, we show that a large domain gap not necessarily leads to a low test accuracy. In this paper, we investigate this problem from the perspective of feature separability empirically and theoretically. Specifically, we propose a dataset-level score based upon feature dispersion to estimate the test accuracy under distribution shift. Our method is inspired by desirable properties of features in representation learning: high inter-class dispersion and high intra-class compactness. Our analysis shows that inter-class dispersion is strongly correlated with the model accuracy, while intra-class compactness does not reflect the generalization performance on OOD data. Extensive experiments demonstrate the superiority of our method in both prediction performance and computational efficiency.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers7
- Energy-based Automated Model EvaluationRu Peng, Heming Zou, Haobo Wang, Yawen Zeng et al.ICLR 2024 · 18 citations
- MaNo: Exploiting Matrix Norm for Unsupervised Accuracy Estimation Under Distribution ShiftsRenchunzi Xie, Ambroise Odonnat, Vasilii Feofanov, Weijian Deng et al.NeurIPS 2024 · 10 citations
- MiraGe: Multimodal Discriminative Representation Learning for Generalizable AI-Generated Image DetectionKuo Shi, Jie Lu, Shanshan Ye, Guangquan Zhang et al.ACM MM 2025 · 3 citations
- A Geometry-Based View of Mahalanobis OOD DetectionDenis Janiak, Jakub Binkowski, Tomasz KajdanowiczICML 2026 · 3 citations
- Suitability Filter: A Statistical Framework for Classifier Evaluation in Real-World Deployment SettingsAngéline Pouget, Mohammad Yaghini, Stephan Rabanser, Nicolas PapernotICML 2025
Builds on16
- WILDS: A Benchmark of in-the-Wild Distribution ShiftsPang Wei Koh, Shiori Sagawa, Henrik Marklund, Sang Michael Xie et al.ICML 2021 · 1,773 citations
- Provable Guarantees for Self-Supervised Deep Learning with Spectral Contrastive LossJeff Z. HaoChen, Colin Wei, Adrien Gaidon, Tengyu MaNeurIPS 2021 · 425 citations
- Mitigating Neural Network Overconfidence with Logit NormalizationHongxin Wei, Renchunzi Xie, Hao Cheng, Lei Feng et al.ICML 2022 · 386 citations
- Learning to Diversify for Single Domain GeneralizationZijian Wang, Yadan Luo, Ruihong Qiu, Zi Huang et al.ICCV 2021 · 339 citations
- Leveraging unlabeled data to predict out-of-distribution performanceSaurabh Garg, Sivaraman Balakrishnan, Zachary Chase Lipton, Behnam Neyshabur et al.ICLR 2022 · 160 citations
Related papers
- WDiscOOD: Out-of-Distribution Detection via Whitened Linear Discriminant AnalysisYiye Chen, Yunzhi Lin, Ruinian Xu, Patricio A. VelaICCV 2023 · 13 citations
- Hierarchical VAEs Know What They Don't KnowJakob Drachmann Havtorn, Jes Frellsen, Søren Hauberg, Lars MaaløeICML 2021 · 87 citations
- Input Complexity and Out-of-distribution Detection with Likelihood-based Generative ModelsJoan Serrà, David Álvarez, Vicenç Gómez, Olga Slizovskaia et al.ICLR 2020 · 307 citations
- Predicting with Confidence on Unseen DistributionsDevin Guillory, Vaishaal Shankar, Sayna Ebrahimi, Trevor Darrell et al.ICCV 2021 · 141 citations
- Quantifying and Improving Transferability in Domain GeneralizationGuojun Zhang, Han Zhao, Yaoliang Yu, Pascal PoupartNeurIPS 2021 · 56 citations
