Ridge Boosting is Both Robust and Efficient
David Bruns-Smith, Zhongming Xie, Avi Feller
Abstract
Estimators in statistics and machine learning must typically trade off between efficiency, having low variance for a fixed target, and distributional robustness, such as multiaccuracy, or having low bias over a range of possible targets. In this paper, we consider a simple estimator, ridge boosting: starting with any initial predictor, perform a single boosting step with (kernel) ridge regression. Surprisingly, we show that ridge boosting simultaneously achieves both efficiency and distributional robustness: for target distribution shifts whose density ratios lie within an RKHS unit ball, this estimator maintains low bias across all such shifts and has variance at the semiparametric efficiency bound for each target. In addition to bridging otherwise distinct research areas, this result has immediate practical value. Since ridge boosting uses only data from the source distribution, researchers can train a single model to obtain both robust and efficient estimates for multiple target estimands at the same time, eliminating the need to fit separate semiparametric efficient estimators for each target. We assess this approach through simulations and an application estimating the age profile of retirement income.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 238bfb3f-a3a2-4379-88c6-04148b05c152Builds on5
- Retiring Adult: New Datasets for Fair Machine LearningFrances Ding, Moritz Hardt, John Miller, Ludwig SchmidtNeurIPS 2021 · 671 citations
- Multicalibration as Boosting for RegressionIra Globus-Harris, Declan Harrison, Michael Kearns, Aaron Roth et al.ICML 2023 · 36 citations
- Causal Isotonic Calibration for Heterogeneous Treatment EffectsLars van der Laan, Ernesto Ulloa-Pérez, Marco Carone, Alex LuedtkeICML 2023 · 18 citations
- Bridging Multicalibration and Out-of-distribution Generalization Beyond Covariate ShiftJiayun Wu, Jiashuo Liu, Peng Cui, Steven WuNeurIPS 2024 · 14 citations
- Kernel Debiased Plug-in Estimation: Simultaneous, Automated Debiasing without Influence Functions for Many Target ParametersBrian M. Cho, Yaroslav Mukhin, Kyra Gan, Ivana MalenicaICML 2024 · 8 citations
Related papers
- Single Point Transductive PredictionNilesh Tripuraneni, Lester MackeyICML 2020 · 5 citations
- Bayesian Nonparametrics Meets Data-Driven Distributionally Robust OptimizationNicola Bariletto, Nhat HoNeurIPS 2024 · 4 citations
- Semi-Supervised Learning with Noisy Proxy Covariates: Generalization Bounds and Distribution RegressionKwangho Kim, Jisu KimICML 2026
- Multiply Robust Estimation for Local Distribution Shifts with Multiple DomainsSteven Wilkins-Reeves, Xu Chen, Qi Ma, Christine Agarwal et al.ICML 2024 · 2 citations
- Estimating Continuous Treatment Effects with Two-Stage Kernel Ridge RegressionSeok-Jin Kim, Kaizheng WangICML 2026 · 1 citation
