Sketched Ridgeless Linear Regression: The Role of Downsampling
Xin Chen, Yicheng Zeng, Siyue Yang, Qiang Sun
摘要
Overparametrization often helps improve the generalization performance. This paper presents a dual view of overparametrization suggesting that downsampling may also help generalize. Focusing on the proportional regime , where represents the sketching size, is the sample size, and is the feature dimensionality, we investigate two out-of-sample prediction risks of the sketched ridgeless least square estimator. Our findings challenge conventional beliefs by showing that downsampling does not always harm generalization but can actually improve it in certain cases. We identify the optimal sketching size that minimizes out-of-sample prediction risks and demonstrate that the optimally sketched estimator exhibits stabler risk curves, eliminating the peaks of those for the full-sample estimator. To facilitate practical implementation, we propose an empirical procedure to determine the optimal sketching size. Finally, we extend our analysis to cover central limit theorems and misspecified models. Numerical studies strongly support our theory.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Asymptotically Free Sketched Ridge Ensembles: Risks, Cross-Validation, and TuningPratik Patil, Daniel LeJeuneICLR 2024 · 被引用 13 次
- High-Dimensional Analysis for Generalized Nonlinear Regression: From Asymptotics to AlgorithmJian Li, Yong Liu, Weiping WangAAAI 2024 · 被引用 4 次
- Implicit Regularization Paths of Weighted Neural RepresentationsJin-Hong Du, Pratik PatilNeurIPS 2024 · 被引用 2 次
- Why Self-Distillation Helps and Hurts: Denoising vs. Signal ForgettingMingqi Wu, Archer Yang, Qiang SunICML 2026
- Prediction Risk and Estimation Risk of the Ridgeless Least Squares Estimator under General Assumptions on Regression ErrorsSungyoon Lee, Sokbae LeeICLR 2025
它引用的顶会 Paper3
- Deep Double Descent: Where Bigger Models and More Data HurtPreetum Nakkiran, Gal Kaplun, Yamini Bansal, Tristan Yang 等ICLR 2020 · 被引用 1,108 次
- Generalization of Two-layer Neural Networks: An Asymptotic ViewpointJimmy Ba, Murat A. Erdogdu, Taiji Suzuki, Denny Wu 等ICLR 2020 · 被引用 77 次
- Asymptotic Normality and Confidence Intervals for Prediction Risk of the Min-Norm Least Squares EstimatorZeng Li, Chuanlong Xie, Qinwen WangICML 2021 · 被引用 4 次
相关 Paper
- Subsample Ridge Ensembles: Equivalences and Generalized Cross-ValidationJin-Hong Du, Pratik Patil, Arun K. KuchibhotlaICML 2023 · 被引用 12 次
- A Fast and Accurate Estimator for Large Scale Linear Model via Data AveragingRui Wang, Yanyan Ouyang, Panpan Yu, Wangli XuNeurIPS 2023 · 被引用 1 次
- Generalized equivalences between subsampling and ridge regularizationPratik Patil, Jin-Hong DuNeurIPS 2023 · 被引用 10 次
- Overfitting Behaviour of Gaussian Kernel Ridgeless Regression: Varying Bandwidth or DimensionalityMarko Medvedev, Gal Vardi, Nati SrebroNeurIPS 2024 · 被引用 9 次
- No Free Lunch from Random Feature Ensembles: Scaling Laws and Near-Optimality ConditionsBenjamin S. Ruben, William Lingxiao Tong, Hamza Tahir Chaudhry, Cengiz PehlevanICML 2025
