Initialization-Dependent Sample Complexity of Linear Predictors and Neural Networks
Roey Magen, Ohad Shamir
摘要
We provide several new results on the sample complexity of vector-valued linear predictors (parameterized by a matrix), and more generally neural networks. Focusing on size-independent bounds, where only the Frobenius norm distance of the parameters from some fixed reference matrix is controlled, we show that the sample complexity behavior can be surprisingly different than what we may expect considering the well-studied setting of scalar-valued linear predictors. This also leads to new sample complexity bounds for feed-forward neural networks, tackling some open questions in the literature, and establishing a new convex linear prediction problem that is provably learnable without uniform convergence.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper3
- Gradient Descent Maximizes the Margin of Homogeneous Neural NetworksKaifeng Lyu, Jian LiICLR 2020 · 被引用 402 次
- The Sample Complexity of One-Hidden-Layer Neural NetworksGal Vardi, Ohad Shamir, Nati SrebroNeurIPS 2022 · 被引用 11 次
- Max-Margin Works while Large Margin Fails: Generalization without Uniform ConvergenceMargalit Glasgow, Colin Wei, Mary Wootters, Tengyu MaICLR 2023
相关 Paper
- Exact Gap between Generalization Error and Uniform Convergence in Random Feature ModelsZitong Yang, Yu Bai, Song MeiICML 2021 · 被引用 19 次
- How many samples are needed to train a deep neural network?Pegah Golestaneh, Mahsa Taheri, Johannes LedererICLR 2025
- On the Optimal Memorization Power of ReLU Neural NetworksGal Vardi, Gilad Yehudai, Ohad ShamirICLR 2022 · 被引用 42 次
- Generalization Analysis of Deep Non-linear Matrix CompletionAntoine Ledent, Rodrigo AlvesICML 2024 · 被引用 5 次
- Fine-grained Generalization Analysis of Vector-Valued LearningLiang Wu, Antoine Ledent, Yunwen Lei, Marius KloftAAAI 2021 · 被引用 11 次
