Support vector machines and linear regression coincide with very high-dimensional features
Navid Ardeshir, Clayton Sanford, Daniel J. Hsu
Abstract
The support vector machine (SVM) and minimum Euclidean norm least squares regression are two fundamentally different approaches to fitting linear models, but they have recently been connected in models for very high-dimensional data through a phenomenon of support vector proliferation, where every training example used to fit an SVM becomes a support vector. In this paper, we explore the generality of this phenomenon and make the following contributions. First, we prove a super-linear lower bound on the dimension (in terms of sample size) required for support vector proliferation in independent feature models, matching the upper bounds from previous works. We further identify a sharp phase transition in Gaussian feature models, bound the width of this transition, and give experimental support for its universality. Finally, we hypothesize that this phase transition occurs only in much higher-dimensional settings in the variant of the SVM, and we present a new geometric characterization of the problem that may elucidate this phenomenon for the general case.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ecdf7f5b-202f-4c72-97ce-ec309c99c111Cited by top-tier papers3
- Benign Overfitting in Multiclass Classification: All Roads Lead to InterpolationKe Wang, Vidya Muthukumar, Christos ThrampoulidisNeurIPS 2021 · 56 citations
- Kernel Memory Networks: A Unifying Framework for Memory ModelingGeorgios Iatropoulos, Johanni Brea, Wulfram GerstnerNeurIPS 2022 · 15 citations
- Finite Smoothing Algorithm for High-Dimensional Support Vector Machines and Quantile RegressionQian Tang, Yikai Zhang, Boxiang WangICML 2024
Builds on1
Related papers
- Deep Principal Support Vector Machines for Nonlinear Sufficient Dimension ReductionYinfeng Chen, Jin Liu, Rui QiuICML 2025
- Overfitting Behaviour of Gaussian Kernel Ridgeless Regression: Varying Bandwidth or DimensionalityMarko Medvedev, Gal Vardi, Nati SrebroNeurIPS 2024 · 9 citations
- A Precise Performance Analysis of Support Vector RegressionHoussem Sifaou, Abla Kammoun, Mohamed-Slim AlouiniICML 2021 · 8 citations
- Near-Tight Margin-Based Generalization Bounds for Support Vector MachinesAllan Grønlund, Lior Kamma, Kasper Green LarsenICML 2020 · 27 citations
- Semi-Supervised Sparse Gaussian Classification: Provable Benefits of Unlabeled DataEyar Azar, Boaz NadlerNeurIPS 2024 · 5 citations
