Understanding the Under-Coverage Bias in Uncertainty Estimation
Yu Bai, Song Mei, Huan Wang, Caiming Xiong
Abstract
Estimating the data uncertainty in regression tasks is often done by learning a quantile function or a prediction interval of the true label conditioned on the input. It is frequently observed that quantile regression -- a vanilla algorithm for learning quantiles with asymptotic guarantees -- tends to under-cover than the desired coverage level in reality. While various fixes have been proposed, a more fundamental understanding of why this under-coverage bias happens in the first place remains elusive. In this paper, we present a rigorous theoretical study on the coverage of uncertainty estimation algorithms in learning quantiles. We prove that quantile regression suffers from an inherent under-coverage bias, in a vanilla setting where we learn a realizable linear quantile function and there is more data than parameters. More quantitatively, for and small , the -quantile learned by quantile regression roughly achieves coverage regardless of the noise distribution, where is the input dimension and is the number of training data. Our theory reveals that this under-coverage bias stems from a certain high-dimensional parameter estimation error that is not implied by existing theories on quantile regression. Experiments on simulated and real data verify our theory and further illustrate the effect of various factors such as sample size and model capacity on the under-coverage bias in more practical setups.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2d86810a-e229-4a26-83ad-a4c99f0cb70bCited by top-tier papers5
- Calibrating Multimodal LearningHuan Ma, Qingyang Zhang, Changqing Zhang, Bingzhe Wu et al.ICML 2023 · 42 citations
- Beyond Confidence: Reliable Models Should Also Consider AtypicalityMert Yüksekgönül, Linjun Zhang, James Y. Zou, Carlos GuestrinNeurIPS 2023 · 32 citations
- Efficient and Differentiable Conformal Prediction with General Function ClassesYu Bai, Song Mei, Huan Wang, Yingbo Zhou et al.ICLR 2022 · 29 citations
- Efficient Uncertainty Quantification and Reduction for Over-Parameterized Neural NetworksZiyi Huang, Henry Lam, Haofeng ZhangNeurIPS 2023 · 21 citations
- Sample-Conditional Coverage in Split-Conformal PredictionJohn C. DuchiNeurIPS 2025 · 2 citations
Builds on4
- Distribution-free binary classification: prediction sets, confidence intervals and calibrationChirag Gupta, Aleksandr Podkopaev, Aaditya RamdasNeurIPS 2020 · 105 citations
- Don't Just Blame Over-parametrization for Over-confidence: Theoretical Analysis of Calibration in Binary ClassificationYu Bai, Song Mei, Huan Wang, Caiming XiongICML 2021 · 47 citations
- Sample Complexity of Uniform Convergence for MulticalibrationEliran Shabat, Lee Cohen, Yishay MansourNeurIPS 2020 · 32 citations
- Uncertainty Sets for Image Classifiers using Conformal PredictionAnastasios Nikolas Angelopoulos, Stephen Bates, Michael I. Jordan, Jitendra MalikICLR 2021 · 31 citations
Related papers
- Improving Conditional Coverage via Orthogonal Quantile RegressionShai Feldman, Stephen Bates, Yaniv RomanoNeurIPS 2021 · 68 citations
- Relaxed Quantile Regression: Prediction Intervals for Asymmetric NoiseThomas Pouplin, Alan Jeffares, Nabeel Seedat, Mihaela van der SchaarICML 2024 · 9 citations
- Asymptotic Normality and Confidence Intervals for Prediction Risk of the Min-Norm Least Squares EstimatorZeng Li, Chuanlong Xie, Qinwen WangICML 2021 · 4 citations
- Conformalized Fairness via Quantile RegressionMeichen Liu, Lei Ding, Dengdeng Yu, Wulong Liu et al.NeurIPS 2022 · 22 citations
- Distribution-Free Model-Agnostic Regression Calibration via Nonparametric MethodsShang Liu, Zhongze Cai, Xiaocheng LiNeurIPS 2023 · 5 citations
