Precise asymptotic generalization for multiclass classification with overparameterized linear models
David Xing Wu, Anant Sahai
摘要
We study the asymptotic generalization of an overparameterized linear model for multiclass classification under the Gaussian covariates bi-level model introduced in Subramanian et al. (2022), where the number of data points, features, and classes all grow together. We fully resolve the conjecture posed in Subramanian et al. ( 2022 ), matching the predicted regimes for generalization. Furthermore, our new lower bounds are akin to an information-theoretic strong converse: they establish that the misclassification rate goes to 0 or 1 asymptotically. One surprising consequence of our tight results is that the min-norm interpolating classifier can be asymptotically suboptimal relative to noninterpolating classifiers in the regime where the min-norm interpolating regressor is known to be optimal. The key to our tight analysis is a new variant of the Hanson-Wright inequality which is broadly useful for multiclass problems with sparse labels. As an application, we show that the same type of analysis can be used to analyze the related multilabel classification problem under the same bi-level ensemble.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Benign Overfitting in Multiclass Classification: All Roads Lead to InterpolationKe Wang, Vidya Muthukumar, Christos ThrampoulidisNeurIPS 2021 · 被引用 56 次
- Implicit Optimization Bias of Next-token Prediction in Linear ModelsChristos ThrampoulidisNeurIPS 2024 · 被引用 19 次
- Provable weak-to-strong generalization via benign overfittingDavid Xing Wu, Anant SahaiICLR 2025
它引用的顶会 Paper4
- More Than a Toy: Random Matrix Models Predict How Real-World Neural Representations GeneralizeAlexander Wei, Wei Hu, Jacob SteinhardtICML 2022 · 被引用 90 次
- Learning Gaussian Mixtures with Generalized Linear Models: Precise Asymptotics in High-dimensionsBruno Loureiro, Gabriele Sicuro, Cédric Gerbelot, Alessandro Pacco 等NeurIPS 2021 · 被引用 70 次
- Benign, Tempered, or Catastrophic: Toward a Refined Taxonomy of OverfittingNeil Mallinar, James B. Simon, Amirhesam Abedsoltan, Parthe Pandit 等NeurIPS 2022 · 被引用 53 次
- Generalization for multiclass classification with overparameterized linear modelsVignesh Subramanian, Rahul Arya, Anant SahaiNeurIPS 2022 · 被引用 12 次
相关 Paper
- Theoretical Insights Into Multiclass Classification: A High-dimensional Asymptotic ViewChristos Thrampoulidis, Samet Oymak, Mahdi SoltanolkotabiNeurIPS 2020 · 被引用 46 次
- Multiclass learning with margin: exponential rates with no bias-variance trade-offStefano Vigogna, Giacomo Meanti, Ernesto De Vito, Lorenzo RosascoICML 2022 · 被引用 3 次
- A Non-Asymptotic Moreau Envelope Theory for High-Dimensional Generalized Linear ModelsLijia Zhou, Frederic Koehler, Pragya Sur, Danica J. Sutherland 等NeurIPS 2022 · 被引用 13 次
- Revisiting Discriminative vs. Generative Classifiers: Theory and ImplicationsChenyu Zheng, Guoqiang Wu, Fan Bao, Yue Cao 等ICML 2023 · 被引用 40 次
- Risk Phase Transitions in Spiked Regression: Alignment Driven Benign and Catastrophic OverfittingJiping Li, Rishi SonthaliaICLR 2026 · 被引用 2 次
