Characterizing Overfitting in Kernel Ridgeless Regression Through the Eigenspectrum
Tin Sum Cheng, Aurélien Lucchi, Anastasis Kratsios, David Belius
摘要
We derive new bounds for the condition number of kernel matrices, which we then use to enhance existing non-asymptotic test error bounds for kernel ridgeless regression (KRR) in the over-parameterized regime for a fixed input dimension. For kernels with polynomial spectral decay, we recover the bound from previous work; for exponential decay, our bound is non-trivial and novel. Our contribution is two-fold: (i) we rigorously prove the phenomena of tempered overfitting and catastrophic overfitting under the sub-Gaussian design assumption, closing an existing gap in the literature; (ii) we identify that the independence of the features plays an important role in guaranteeing tempered overfitting, raising concerns about approximating KRR generalization using the Gaussian design assumption in previous literature. Contributions This work aims to understand the overfitting behaviour for kernel ridge regression (KRR). Our main contributions are summarized as follows: 1. We rigorously derive tight non-asymptotic upper and lower bounds for the test error of the minimum norm interpolant under a sub-Gaussian design assumption on the features. While this assumption assumes independence of the features, this point will be relaxed in contribution #3.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- A Comprehensive Analysis on the Learning Curve in Kernel Ridge RegressionTin Sum Cheng, Aurélien Lucchi, Anastasis Kratsios, David BeliusNeurIPS 2024 · 被引用 7 次
- Beyond Benign Overfitting in Nadaraya-Watson InterpolatorsDaniel Barzilai, Guy Kornowski, Ohad ShamirNeurIPS 2025 · 被引用 2 次
- Harmful Overfitting in Sobolev SpacesKedar Karhadkar, Alexander Sietsema, Deanna Needell, Guido MontufarICML 2026
- Kernel Regression in Structured Non-IID Settings: Theory and Implications for Denoising Score LearningDechen Zhang, Zhenmei Shi, Yi Zhang, Yingyu Liang 等NeurIPS 2025
- Training-Free Determination of Network Width via Neural Tangent KernelTatsumi Sunada, Toshihiko Yamasaki, Atsuto MakiICLR 2026
它引用的顶会 Paper8
- Spectrum Dependent Learning Curves in Kernel Regression and Wide Neural NetworksBlake Bordelon, Abdulkadir Canatar, Cengiz PehlevanICML 2020 · 被引用 245 次
- Learning curves of generic features maps for realistic datasets with a teacher-student modelBruno Loureiro, Cédric Gerbelot, Hugo Cui, Sebastian Goldt 等NeurIPS 2021 · 被引用 170 次
- Generalization Error Rates in Kernel Regression: The Crossover from the Noiseless to Noisy RegimeHugo Cui, Bruno Loureiro, Florent Krzakala, Lenka ZdeborováNeurIPS 2021 · 被引用 109 次
- Implicit Regularization of Random Feature ModelsArthur Jacot, Berfin Simsek, Francesco Spadaro, Clément Hongler 等ICML 2020 · 被引用 83 次
- Mind the spikes: Benign overfitting of kernels and neural networks in fixed dimensionMoritz Haas, David Holzmüller, Ulrike von Luxburg, Ingo SteinwartNeurIPS 2023 · 被引用 30 次
相关 Paper
- An Agnostic View on the Cost of Overfitting in (Kernel) Ridge RegressionLijia Zhou, James B. Simon, Gal Vardi, Nathan SrebroICLR 2024 · 被引用 2 次
- Overfitting Behaviour of Gaussian Kernel Ridgeless Regression: Varying Bandwidth or DimensionalityMarko Medvedev, Gal Vardi, Nati SrebroNeurIPS 2024 · 被引用 9 次
- Benign, Tempered, or Catastrophic: Toward a Refined Taxonomy of OverfittingNeil Mallinar, James B. Simon, Amirhesam Abedsoltan, Parthe Pandit 等NeurIPS 2022 · 被引用 53 次
- Towards Theoretical Understanding of Learning Large-scale Dependent Data via Random FeaturesChao Wang, Xin Bing, Xin He, Caixing WangICML 2024 · 被引用 1 次
- Target alignment in truncated kernel ridge regressionArash A. Amini, Richard Baumgartner, Dai FengNeurIPS 2022 · 被引用 4 次
