Overcoming the curse of dimensionality with Laplacian regularization in semi-supervised learning
Vivien Cabannes, Loucas Pillaud-Vivien, Francis R. Bach, Alessandro Rudi
Abstract
As annotations of data can be scarce in large-scale practical problems, leveraging unlabelled examples is one of the most important aspects of machine learning. This is the aim of semi-supervised learning. To benefit from the access to unlabelled data, it is natural to diffuse smoothly knowledge of labelled data to unlabelled one. This induces to the use of Laplacian regularization. Yet, current implementations of Laplacian regularization suffer from several drawbacks, notably the well-known curse of dimensionality. In this paper, we provide a statistical analysis to overcome those issues, and unveil a large body of spectral filtering methods that exhibit desirable behaviors. They are implemented through (reproducing) kernel methods, for which we provide realistic computational guidelines in order to make our method usable with large amounts of data.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 95990274-ff5e-4a33-9f46-cab2b0ade609Cited by top-tier papers6
- The SSL Interplay: Augmentations, Inductive Bias, and GeneralizationVivien Cabannes, Bobak Toussi Kiani, Randall Balestriero, Yann LeCun et al.ICML 2023 · 43 citations
- Sobolev Acceleration and Statistical Optimality for Learning Elliptic Equations via Gradient DescentYiping Lu, José H. Blanchet, Lexing YingNeurIPS 2022 · 15 citations
- Learning Globally Smooth Functions on ManifoldsJuan Cerviño, Luiz F. O. Chamon, Benjamin David Haeffele, René Vidal et al.ICML 2023 · 6 citations
- Active Labeling: Streaming Stochastic GradientsVivien Cabannes, Francis R. Bach, Vianney Perchet, Alessandro RudiNeurIPS 2022 · 2 citations
- Fast Algorithms for Hypergraph PageRank with Applications to Semi-Supervised LearningKonstantinos Ameranis, Adela Frances DePavia, Lorenzo Orecchia, Erasmo TaniICML 2024 · 1 citation
Builds on4
- HowTo100M: Learning a Text-Video Embedding by Watching Hundred Million Narrated Video ClipsAntoine Miech, Dimitri Zhukov, Jean-Baptiste Alayrac, Makarand Tapaswi et al.ICCV 2019 · 1,437 citations
- Structured Prediction with Partial Labelling through the Infimum LossVivien Cabannes, Alessandro Rudi, Francis R. BachICML 2020 · 50 citations
- Bayes Consistency vs. H-Consistency: The Interplay between Surrogate Loss Functions and the Scoring Function ClassMingyuan Zhang, Shivani AgarwalNeurIPS 2020 · 42 citations
- Deep Neural Tangent Kernel and Laplace Kernel Have the Same RKHSLin Chen, Sheng XuICLR 2021 · 3 citations
Related papers
- Semi-supervised Conditional Density Estimation with Wasserstein Laplacian RegularisationOlivier Graffeuille, Yun Sing Koh, Jörg Wicker, Moritz K. LehmannAAAI 2022 · 4 citations
- Spectrally Transformed Kernel RegressionRuntian Zhai, Rattana Pukdee, Roger Jin, Maria-Florina Balcan et al.ICLR 2024 · 3 citations
- Label Propagation with Weak SupervisionRattana Pukdee, Dylan Sam, Pradeep Kumar Ravikumar, Nina BalcanICLR 2023
- Supervised and Semi-Supervised Diffusion Maps with Label-Driven DiffusionHarel Mendelman, Ronen TalmonICLR 2025
- Functional Regularization for Representation Learning: A Unified Theoretical PerspectiveSiddhant Garg, Yingyu LiangNeurIPS 2020 · 27 citations
