The Effect of the Intrinsic Dimension on the Generalization of Quadratic Classifiers
Fabian Latorre, Leello Tadesse Dadi, Paul Rolland, Volkan Cevher
Abstract
It has been recently observed that neural networks, unlike kernel methods, enjoy a reduced sample complexity when the distribution is isotropic (i.e., when the covariance matrix is the identity). We find that this sensitivity to the data distribution is not exclusive to neural networks, and the same phenomenon can be observed on the class of quadratic classifiers (i.e., the sign of a quadratic polynomial) with a nuclear-norm constraint. We demonstrate this by deriving an upper bound on the Rademacher Complexity that depends on two key quantities: (i) the intrinsic dimension, which is a measure of isotropy, and (ii) the largest eigenvalue of the second moment (covariance) matrix of the distribution. Our result improves the dependence on the dimension over the best previously known bound and precisely quantifies the relation between the sample complexity and the level of isotropy of the distribution.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers3
- Light Schrödinger BridgeAlexander Korotin, Nikita Gushchin, Evgeny BurnaevICLR 2024 · 16 citations
- A Statistical Learning Perspective on Semi-dual Adversarial Neural Optimal Transport SolversRoman Tarasov, Petr Mokrov, Milena Gazdieva, Evgeny Burnaev et al.ICLR 2026 · 2 citations
- Local Intrinsic Dimensional EntropyRohan Ghosh, Mehul MotaniAAAI 2023 · 2 citations
Builds on5
- Fantastic Generalization Measures and Where to Find ThemYiding Jiang, Behnam Neyshabur, Hossein Mobahi, Dilip Krishnan et al.ICLR 2020 · 705 citations
- The Intrinsic Dimension of Images and Its Impact on LearningPhillip Pope, Chen Zhu, Ahmed Abdelkader, Micah Goldblum et al.ICLR 2021 · 381 citations
- When Do Neural Networks Outperform Kernel Methods?Behrooz Ghorbani, Song Mei, Theodor Misiakiewicz, Andrea MontanariNeurIPS 2020 · 217 citations
- Multiplicative Interactions and Where to Find ThemSiddhant M. Jayakumar, Wojciech M. Czarnecki, Jacob Menick, Jonathan Schwarz et al.ICLR 2020 · 152 citations
- Beyond Linearization: On Quadratic and Higher-Order Approximation of Wide Neural NetworksYu Bai, Jason D. LeeICLR 2020 · 128 citations
Related papers
- The Effect of Manifold Entanglement and Intrinsic Dimensionality on LearningDaniel Kienitz, Ekaterina Komendantskaya, Michael A. LonesAAAI 2022 · 6 citations
- Why Are Convolutional Nets More Sample-Efficient than Fully-Connected Nets?Zhiyuan Li, Yi Zhang, Sanjeev AroraICLR 2021 · 22 citations
- Why High-rank Neural Networks Generalize?: An Algebraic Framework with RKHSsYuka Hashimoto, Sho Sonoda, Isao Ishikawa, Masahiro IkedaICLR 2026 · 1 citation
- Gradient-Based Feature Learning under Structured DataAlireza Mousavi-Hosseini, Denny Wu, Taiji Suzuki, Murat A. ErdogduNeurIPS 2023 · 36 citations
- Dropout: Explicit Forms and Capacity ControlRaman Arora, Peter L. Bartlett, Poorya Mianjy, Nathan SrebroICML 2021 · 43 citations
