Learnability of high-dimensional targets by two-parameter models and gradient flow
Dmitry Yarotsky
摘要
We explore the theoretical possibility of learning -dimensional targets with -parameter models by gradient flow (GF) when . Our main result shows that if the targets are described by a particular -dimensional probability distribution, then there exist models with as few as two parameters that can learn the targets with arbitrarily high success probability. On the other hand, we show that for there is necessarily a large subset of GF-non-learnable targets. In particular, the set of learnable targets is not dense in , and any subset of homeomorphic to the -dimensional sphere contains non-learnable targets. Finally, we observe that the model in our main theorem on almost guaranteed two-parameter learning is constructed using a hierarchical procedure and as a result is not expressible by a single elementary function. We show that this limitation is essential in the sense that most models written in terms of elementary functions cannot achieve the learnability demonstrated in this theorem.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper3
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Elementary superexpressive activationsDmitry YarotskyICML 2021 · 被引用 46 次
- Constructive Universal High-Dimensional Distribution Generation through Deep ReLU NetworksDmytro Perekrestenko, Stephan Müller, Helmut BölcskeiICML 2020 · 被引用 15 次
相关 Paper
- Mean-Field Analysis for Learning Subspace-Sparse Polynomials with Gaussian InputZiang Chen, Rong GeNeurIPS 2024 · 被引用 1 次
- The Computational Advantage of Depth in Learning High-Dimensional Hierarchical TargetsYatin Dandi, Luca Pesce, Lenka Zdeborová, Florent KrzakalaNeurIPS 2025 · 被引用 3 次
- Neural Networks Learn Generic Multi-Index Models Near Information-Theoretic LimitBohan Zhang, Zihao Wang, Hengyu Fu, Jason D. LeeICLR 2026 · 被引用 3 次
- A solvable model of learning generative diffusion: theory and insightsHugo Cui, Cengiz Pehlevan, Yue M. LuNeurIPS 2025 · 被引用 11 次
- Deep Network Approximation in Terms of Intrinsic ParametersZuowei Shen, Haizhao Yang, Shijun ZhangICML 2022 · 被引用 13 次
