Learning Prediction Intervals for Model Performance
Benjamin Elder, Matthew Arnold, Anupama Murthi, Jirí Navrátil
Abstract
Understanding model performance on unlabeled data is a fundamental challenge of developing, deploying, and maintaining AI systems. Model performance is typically evaluated using test sets or periodic manual quality assessments, both of which require laborious manual data labeling. Automated performance prediction techniques aim to mitigate this burden, but potential inaccuracy and a lack of trust in their predictions has prevented their widespread adoption. We address this core problem of performance prediction uncertainty with a method to compute prediction intervals for model performance. Our methodology uses transfer learning to train an uncertainty model to estimate the uncertainty of model performance predictions. We evaluate our approach across a wide range of drift conditions and show substantial improvement over competitive baselines. We believe this result makes prediction intervals, and performance prediction in general, significantly more practical for real-world use.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8a14cc54-58c2-4eb8-95f3-21471222343dCited by top-tier papers2
- Post-hoc Uncertainty Learning Using a Dirichlet Meta-ModelMaohao Shen, Yuheng Bu, Prasanna Sattigeri, Soumya Ghosh et al.AAAI 2023 · 51 citations
- SurvUnc: A Meta-Model Based Uncertainty Quantification Framework for Survival AnalysisYu Liu, Weiyao Tao, Tong Xia, Simon Knight et al.KDD 2025 · 3 citations
Related papers
- Optimal Aggregation of Prediction Intervals under Unsupervised Domain ShiftJiawei Ge, Debarghya Mukherjee, Jianqing FanNeurIPS 2024 · 6 citations
- Are Labels Always Necessary for Classifier Accuracy Evaluation?Weijian Deng, Liang ZhengCVPR 2021
- Characterizing Out-of-Distribution Error via Optimal TransportYuzhe Lu, Yilong Qin, Runtian Zhai, Andrew Shen et al.NeurIPS 2023 · 22 citations
- Estimating Model Performance Under Covariate Shift Without LabelsJakub Bialek, Juhani Kivimäki, Wojtek Kuberski, Nikolaos PerrakisNeurIPS 2025 · 10 citations
- Unlabelled Data Improves Bayesian Uncertainty Calibration under Covariate ShiftAlex J. Chan, Ahmed M. Alaa, Zhaozhi Qian, Mihaela van der SchaarICML 2020 · 42 citations
