Stability of Random Forests and Coverage of Random-Forest Prediction Intervals
Yan Wang, Huaiqing Wu, Dan Nettleton
Abstract
We establish stability of random forests under the mild condition that the squared response () does not have a heavy tail. In particular, our analysis holds for the practical version of random forests that is implemented in popular packages like randomForest in R. Empirical results show that stability may persist even beyond our assumption and hold for heavy-tailed . Using the stability property, we prove a non-asymptotic lower bound for the coverage probability of prediction intervals constructed from the out-of-bag error of random forests. With another mild condition that is typically satisfied when is continuous, we also establish a complementary upper bound, which can be similarly established for the jackknife prediction interval constructed from an arbitrary stable algorithm. We also discuss the asymptotic coverage probability under assumptions weaker than those considered in previous literature. Our work implies that random forests, with its stability property, is an effective machine learning method that can provide not only satisfactory point prediction but also justified interval prediction at almost no extra computational cost.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ef5e7f9d-f1b3-47d8-a7df-b8d48e2b409cCited by top-tier papers4
- Leave-One-Out Stable Conformal PredictionKiljae Lee, Yuan ZhangICLR 2025
- Tight Asymptotics of Extreme Order StatisticsJosé Correa, Frederik Mallmann-Trenn, Matías RomeroNeurIPS 2025
- Generalization Bounds for Model-based Algorithm ConfigurationZhiyang Chen, Hailong Yao, Xia YinNeurIPS 2025
- Feature Bagging Provides StabilityYuheng Ma, Qiang SunICML 2026
Builds on2
Related papers
- Asymptotics of the Bootstrap via Stability with Applications to Inference with Model SelectionMorgane Austern, Vasilis SyrgkanisNeurIPS 2021 · 1 citation
- Cross-validation Confidence Intervals for Test ErrorPierre Bayle, Alexandre Bayle, Lucas Janson, Lester MackeyNeurIPS 2020 · 76 citations
- Discriminative Jackknife: Quantifying Uncertainty in Deep Learning via Higher-Order Influence FunctionsAhmed M. Alaa, Mihaela van der SchaarICML 2020 · 59 citations
- Understanding the Under-Coverage Bias in Uncertainty EstimationYu Bai, Song Mei, Huan Wang, Caiming XiongNeurIPS 2021 · 18 citations
- Conformal Prediction using Conditional HistogramsMatteo Sesia, Yaniv RomanoNeurIPS 2021 · 106 citations
