Uniform Convergence of Interpolators: Gaussian Width, Norm Bounds and Benign Overfitting
Frederic Koehler, Lijia Zhou, Danica J. Sutherland, Nathan Srebro
摘要
We consider interpolation learning in high-dimensional linear regression with Gaussian data, and prove a generic uniform convergence guarantee on the generalization error of interpolators in an arbitrary hypothesis class in terms of the class's Gaussian width. Applying the generic bound to Euclidean norm balls recovers the consistency result of Bartlett et al. (2020) for minimum-norm interpolators, and confirms a prediction of Zhou et al. ( 2020 ) for near-minimal-norm interpolators in the special case of Gaussian data. We demonstrate the generality of the bound by applying it to the simplex, obtaining a novel consistency result for minimum 1 -norm interpolators (basis pursuit). Our results show how norm-based generalization bounds can explain and be used to analyze benign overfitting, at least in some settings. * These authors contributed equally. 1 Negrea et al. (2020) argue that Bartlett et al. (2020)'s proof technique is fundamentally based on uniform convergence of a surrogate predictor; Yang et al. (2021) study a closely related setting with a uniform convergence-type argument, but do not establish consistency. We discuss both papers in more detail in Section 4. 35th Conference on Neural Information Processing Systems (NeurIPS 2021).
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper32
- Towards Last-layer Retraining for Group Robustness with Fewer AnnotationsTyler LaBonte, Vidya Muthukumar, Abhishek KumarNeurIPS 2023 · 被引用 73 次
- Benign, Tempered, or Catastrophic: Toward a Refined Taxonomy of OverfittingNeil Mallinar, James B. Simon, Amirhesam Abedsoltan, Parthe Pandit 等NeurIPS 2022 · 被引用 53 次
- Predicting Out-of-Distribution Error with the Projection NormYaodong Yu, Zitong Yang, Alexander Wei, Yi Ma 等ICML 2022 · 被引用 51 次
- Fast rates for noisy interpolation require rethinking the effect of inductive biasKonstantin Donhauser, Nicolò Ruggeri, Stefan Stojanovic, Fanny YangICML 2022 · 被引用 24 次
- Regularization properties of adversarially-trained linear regressionAntônio H. Ribeiro, Dave Zachariah, Francis R. Bach, Thomas B. SchönNeurIPS 2023 · 被引用 23 次
它引用的顶会 Paper2
相关 Paper
- Minimum Norm Interpolation Meets The Local Theory of Banach SpacesGil Kur, Pedro Abdalla, Pierre Bizeul, Fanny YangICML 2024 · 被引用 3 次
- A Non-Asymptotic Moreau Envelope Theory for High-Dimensional Generalized Linear ModelsLijia Zhou, Frederic Koehler, Pragya Sur, Danica J. Sutherland 等NeurIPS 2022 · 被引用 13 次
- Uniform Convergence with Square-Root Lipschitz LossLijia Zhou, Zhen Dai, Frederic Koehler, Nati SrebroNeurIPS 2023 · 被引用 2 次
- Overfitting Behaviour of Gaussian Kernel Ridgeless Regression: Varying Bandwidth or DimensionalityMarko Medvedev, Gal Vardi, Nati SrebroNeurIPS 2024 · 被引用 9 次
- Universal Consistency of Wide and Deep ReLU Neural Networks and Minimax Optimal Convergence Rates for Kolmogorov-Donoho Optimal Function ClassesHyunouk Ko, Xiaoming HuoICML 2024 · 被引用 1 次
