On Uniform Convergence and Low-Norm Interpolation Learning
Lijia Zhou, Danica J. Sutherland, Nati Srebro
Abstract
We consider an underdetermined noisy linear regression model where the minimum-norm interpolating predictor is known to be consistent, and ask: can uniform convergence in a norm ball, or at least (following Nagarajan and Kolter) the subset of a norm ball that the algorithm selects on a typical input set, explain this success? We show that uniformly bounding the difference between empirical and population errors cannot show any learning in the norm ball, and cannot show consistency for any set, even one depending on the exact algorithm and distribution. But we argue we can explain the consistency of the minimal-norm interpolator with a slightly weaker, yet standard, notion: uniform convergence of zero-error predictors in a norm ball. We use this to bound the generalization error of low-(but not minimal-) norm interpolating predictors.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext fbf16de2-b463-4c6d-9165-c4c8f2fb173bCited by top-tier papers15
- Assessing Generalization of SGD via DisagreementYiding Jiang, Vaishnavh Nagarajan, Christina Baek, J. Zico KolterICLR 2022 · 134 citations
- Agreement-on-the-line: Predicting the Performance of Neural Networks under Distribution ShiftChristina Baek, Yiding Jiang, Aditi Raghunathan, J. Zico KolterNeurIPS 2022 · 120 citations
- Towards Last-layer Retraining for Group Robustness with Fewer AnnotationsTyler LaBonte, Vidya Muthukumar, Abhishek KumarNeurIPS 2023 · 73 citations
- On Linear Stability of SGD and Input-Smoothness of Neural NetworksChao Ma, Lexing YingNeurIPS 2021 · 73 citations
- Uniform Convergence of Interpolators: Gaussian Width, Norm Bounds and Benign OverfittingFrederic Koehler, Lijia Zhou, Danica J. Sutherland, Nathan SrebroNeurIPS 2021 · 65 citations
Builds on4
- Deep Double Descent: Where Bigger Models and More Data HurtPreetum Nakkiran, Gal Kaplun, Yamini Bansal, Tristan Yang et al.ICLR 2020 · 1,108 citations
- Exact expressions for double descent and implicit regularization via surrogate random designMichal Derezinski, Feynman T. Liang, Michael W. MahoneyNeurIPS 2020 · 81 citations
- Generalization of Two-layer Neural Networks: An Asymptotic ViewpointJimmy Ba, Murat A. Erdogdu, Taiji Suzuki, Denny Wu et al.ICLR 2020 · 77 citations
- In Defense of Uniform Convergence: Generalization via Derandomization with an Application to Interpolating PredictorsJeffrey Negrea, Gintare Karolina Dziugaite, Daniel M. RoyICML 2020 · 66 citations
Related papers
- Exact Gap between Generalization Error and Uniform Convergence in Random Feature ModelsZitong Yang, Yu Bai, Song MeiICML 2021 · 19 citations
- Minimum Norm Interpolation Meets The Local Theory of Banach SpacesGil Kur, Pedro Abdalla, Pierre Bizeul, Fanny YangICML 2024 · 3 citations
- Fast rates for noisy interpolation require rethinking the effect of inductive biasKonstantin Donhauser, Nicolò Ruggeri, Stefan Stojanovic, Fanny YangICML 2022 · 24 citations
- A Non-Asymptotic Moreau Envelope Theory for High-Dimensional Generalized Linear ModelsLijia Zhou, Frederic Koehler, Pragya Sur, Danica J. Sutherland et al.NeurIPS 2022 · 13 citations
- Overfitting Behaviour of Gaussian Kernel Ridgeless Regression: Varying Bandwidth or DimensionalityMarko Medvedev, Gal Vardi, Nati SrebroNeurIPS 2024 · 9 citations
