Semi-Supervised Learning with Noisy Proxy Covariates: Generalization Bounds and Distribution Regression
Kwangho Kim, Jisu Kim
Abstract
In many modern machine learning pipelines, abundant pretrained representations serve as noisy proxy covariates, while task-specific labels remain scarce. We study semi-supervised regression in this setting, and propose a simple two stage estimator that learns kernel eigenfeatures from all proxy covariates and fits a ridge predictor on labeled data. We derive finite sample bounds showing that fast labeled sample rates are recovered when proxy perturbation is controlled and unlabeled proxy covariates are sufficiently abundant. We also show that distribution regression is a direct special case, with analogous guarantees when the finite bag size is large enough. Experiments show consistent gains over supervised and semi-supervised baselines, especially in low label regimes.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d790fe67-e37e-443f-aab4-1c6d2071947eBuilds on2
- Overcoming the curse of dimensionality with Laplacian regularization in semi-supervised learningVivien Cabannes, Loucas Pillaud-Vivien, Francis R. Bach, Alessandro RudiNeurIPS 2021 · 23 citations
- Learning to Embed Distributions via Maximum Kernel EntropyOleksii Kachaiev, Stefano RecanatesiNeurIPS 2024 · 3 citations
Related papers
- Semi-supervised Active Linear RegressionNived Rajaraman, Devvrit, Pranjal AwasthiNeurIPS 2022 · 1 citation
- High-dimensional Analysis of Knowledge Distillation: Weak-to-Strong Generalization and Scaling LawsMuhammed Emrullah Ildiz, Halil Alperen Gozeten, Ege Onur Taga, Marco Mondelli et al.ICLR 2025
- Ridge Boosting is Both Robust and EfficientDavid Bruns-Smith, Zhongming Xie, Avi FellerNeurIPS 2025
- Improved Scaling Laws via Weak-to-Strong Generalization in Random Features Ridge RegressionDiyuan Wu, Lehan Chen, Theodor Misiakiewicz, Marco MondelliICML 2026
- The Power and Limitation of Pretraining-Finetuning for Linear Regression under Covariate ShiftJingfeng Wu, Difan Zou, Vladimir Braverman, Quanquan Gu et al.NeurIPS 2022 · 29 citations
