High-Dimensional Kernel Methods under Covariate Shift: Data-Dependent Implicit Regularization
Yihang Chen, Fanghui Liu, Taiji Suzuki, Volkan Cevher
Abstract
This paper studies kernel ridge regression in high dimensions under covariate shifts and analyzes the role of importance re-weighting. We first derive the asymptotic expansion of high dimensional kernels under covariate shifts. By a bias-variance decomposition, we theoretically demonstrate that the re-weighting strategy allows for decreasing the variance. For bias, we analyze the regularization of the arbitrary or well-chosen scale, showing that the bias can behave very differently under different regularization scales. In our analysis, the bias and variance can be characterized by the spectral decay of a data-dependent regularized kernel: the original kernel matrix associated with an additional re-weighting matrix, and thus the re-weighting strategy can be regarded as a data-dependent regularization for better understanding. Besides, our analysis provides asymptotic expansion of kernel functions/vectors under covariate shift, which has its own interest.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b0607db6-5dc4-4fd5-9799-157f46308b63Cited by top-tier papers4
- Transfer Learning for Benign Overfitting in High-Dimensional Linear RegressionYeichan Kim, Ilmun Kim, Seyoung ParkNeurIPS 2025 · 2 citations
- Mixed-Sample SGD: an End-to-end Analysis of Supervised Transfer LearningYuyang Deng, Samory KpotufeNeurIPS 2025 · 1 citation
- Benign Overfitting in Out-of-Distribution Generalization of Linear ModelsShange Tang, Jiayun Wu, Jianqing Fan, Chi JinICLR 2025
- Kernel Regression in Structured Non-IID Settings: Theory and Implications for Denoising Score LearningDechen Zhang, Zhenmei Shi, Yi Zhang, Yingyu Liang et al.NeurIPS 2025
Builds on14
- When Do Neural Networks Outperform Kernel Methods?Behrooz Ghorbani, Song Mei, Theodor Misiakiewicz, Andrea MontanariNeurIPS 2020 · 217 citations
- Rethinking Importance Weighting for Deep Learning under Distribution ShiftTongtong Fang, Nan Lu, Gang Niu, Masashi SugiyamaNeurIPS 2020 · 179 citations
- Optimal Regularization can Mitigate Double DescentPreetum Nakkiran, Prayaag Venkat, Sham M. Kakade, Tengyu MaICLR 2021 · 148 citations
- Kernel Alignment Risk Estimator: Risk Prediction from Training DataArthur Jacot, Berfin Simsek, Francesco Spadaro, Clément Hongler et al.NeurIPS 2020 · 74 citations
- Overparameterization Improves Robustness to Covariate Shift in High DimensionsNilesh Tripuraneni, Ben Adlam, Jeffrey PenningtonNeurIPS 2021 · 50 citations
Related papers
- Computational Efficiency under Covariate Shift in Kernel Ridge RegressionAndrea Della Vecchia, Arnaud Mavakala Watusadisi, Ernesto De Vito, Lorenzo RosascoNeurIPS 2025 · 5 citations
- Optimal Ridge Regularization for Out-of-Distribution PredictionPratik Patil, Jin-Hong Du, Ryan J. TibshiraniICML 2024 · 23 citations
- Optimal Regularization for Performative LearningEdwige Cyffers, Alireza Mirrokni, Marco MondelliICML 2026 · 1 citation
- Estimating Continuous Treatment Effects with Two-Stage Kernel Ridge RegressionSeok-Jin Kim, Kaizheng WangICML 2026 · 1 citation
- A Comprehensive Analysis on the Learning Curve in Kernel Ridge RegressionTin Sum Cheng, Aurélien Lucchi, Anastasis Kratsios, David BeliusNeurIPS 2024 · 7 citations
