Initialization-Dependent Sample Complexity of Linear Predictors and Neural Networks
Roey Magen, Ohad Shamir
Abstract
We provide several new results on the sample complexity of vector-valued linear predictors (parameterized by a matrix), and more generally neural networks. Focusing on size-independent bounds, where only the Frobenius norm distance of the parameters from some fixed reference matrix is controlled, we show that the sample complexity behavior can be surprisingly different than what we may expect considering the well-studied setting of scalar-valued linear predictors. This also leads to new sample complexity bounds for feed-forward neural networks, tackling some open questions in the literature, and establishing a new convex linear prediction problem that is provably learnable without uniform convergence.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ca18a892-a2b0-4334-8f24-8daa9078d084Cited by top-tier papers1
Ask how each one uses itBuilds on3
- Gradient Descent Maximizes the Margin of Homogeneous Neural NetworksKaifeng Lyu, Jian LiICLR 2020 · 402 citations
- The Sample Complexity of One-Hidden-Layer Neural NetworksGal Vardi, Ohad Shamir, Nati SrebroNeurIPS 2022 · 11 citations
- Max-Margin Works while Large Margin Fails: Generalization without Uniform ConvergenceMargalit Glasgow, Colin Wei, Mary Wootters, Tengyu MaICLR 2023
Related papers
- Exact Gap between Generalization Error and Uniform Convergence in Random Feature ModelsZitong Yang, Yu Bai, Song MeiICML 2021 · 19 citations
- How many samples are needed to train a deep neural network?Pegah Golestaneh, Mahsa Taheri, Johannes LedererICLR 2025
- On the Optimal Memorization Power of ReLU Neural NetworksGal Vardi, Gilad Yehudai, Ohad ShamirICLR 2022 · 42 citations
- Generalization Analysis of Deep Non-linear Matrix CompletionAntoine Ledent, Rodrigo AlvesICML 2024 · 5 citations
- Fine-grained Generalization Analysis of Vector-Valued LearningLiang Wu, Antoine Ledent, Yunwen Lei, Marius KloftAAAI 2021 · 11 citations
