Bootstrap in High Dimension with Low Computation
Henry Lam, Zhenyuan Liu
Abstract
The bootstrap is a popular data-driven method to quantify statistical uncertainty, but for modern high-dimensional problems, it could suffer from huge computational costs due to the need to repeatedly generate resamples and refit models. We study the use of bootstraps in high-dimensional environments with a small number of resamples. In particular, we show that with a recent "cheap" bootstrap perspective, using a number of resamples as small as one could attain valid coverage even when the dimension grows closely with the sample size, thus strongly supporting the implementability of the bootstrap for large-scale problems. We validate our theoretical results and compare the performance of our approach with other benchmarks via a range of experiments.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ed25dd3d-e4b4-4186-aef4-01d45c14a531Builds on1
Related papers
- Centroid Approximation for Bootstrap: Improving Particle Quality at InferenceMao Ye, Qiang LiuICML 2022
- Orthogonal Bootstrap: Efficient Simulation of Input UncertaintyKaizhao Liu, José H. Blanchet, Lexing Ying, Yiping LuICML 2024 · 2 citations
- Error Estimation for Sketched SVD via the BootstrapMiles E. Lopes, N. Benjamin Erichson, Michael W. MahoneyICML 2020 · 12 citations
- Simultaneous Inference for Massive Data: Distributed BootstrapYang Yu, Shih-Kang Chao, Guang ChengICML 2020 · 17 citations
- Neural BootstrapperMinsuk Shin, Hyungjoo Cho, Hyun-seok Min, Sungbin LimNeurIPS 2021 · 10 citations
