Subsampled Ensemble Can Improve Generalization Tail Exponentially
Huajie Qian, Donghao Ying, Henry Lam, Wotao Yin
Abstract
Ensemble learning is a popular technique to improve the accuracy of machine learning models. It traditionally hinges on the rationale that aggregating multiple weak models can lead to better models with lower variance and hence higher stability, especially for discontinuous base learners. In this paper, we provide a new perspective on ensembling. By selecting the most frequently generated model from the base learner when repeatedly applied to subsamples, we can attain exponentially decaying tails for the excess risk, even if the base learner suffers from slow (i.e., polynomial) decay rates. This tail enhancement power of ensembling applies to base learners that have reasonable predictive power to begin with and is stronger than variance reduction in the sense of exhibiting rate improvement. We demonstrate how our ensemble methods can substantially improve out-of-sample performances in a range of numerical examples involving heavy-tailed data or intrinsically slow rates.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on10
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Why are Adaptive Methods Good for Attention Models?Jingzhao Zhang, Sai Praneeth Karimireddy, Andreas Veit, Seungyeon Kim et al.NeurIPS 2020 · 397 citations
- Reinforcement Learning for Integer Programming: Learning to CutYunhao Tang, Shipra Agrawal, Yuri FaenzaICML 2020 · 224 citations
- Stochastic Optimization with Heavy-Tailed Noise via Accelerated Gradient ClippingEduard Gorbunov, Marina Danilova, Alexander V. GasnikovNeurIPS 2020 · 181 citations
- High-probability Bounds for Non-Convex Stochastic Optimization with Heavy TailsAshok Cutkosky, Harsh MehtaNeurIPS 2021 · 119 citations
Related papers
- Theoretical Guarantees of Learning Ensembling Strategies with Applications to Time Series ForecastingHilaf Hasson, Danielle C. Maddix, Bernie Wang, Gaurav Gupta et al.ICML 2023 · 4 citations
- United We Stand: Using Epoch-Wise Agreement of Ensembles to Combat OverfitUri Stern, Daniel Shwartz, Daphna WeinshallAAAI 2024
- When are ensembles really effective?Ryan Theisen, Hyunsuk Kim, Yaoqing Yang, Liam Hodgkinson et al.NeurIPS 2023 · 29 citations
- Predictive inference is free with the jackknife+-after-bootstrapByol Kim, Chen Xu, Rina Foygel BarberNeurIPS 2020 · 105 citations
- Automatic Unsupervised Ensemble Outlier Model SelectionHong-Phuc Phan, Tuan-Anh Vu, Tung Kieu, Sơn Hà Xuân et al.ICML 2026
