Learnability of high-dimensional targets by two-parameter models and gradient flow
Dmitry Yarotsky
Abstract
We explore the theoretical possibility of learning -dimensional targets with -parameter models by gradient flow (GF) when . Our main result shows that if the targets are described by a particular -dimensional probability distribution, then there exist models with as few as two parameters that can learn the targets with arbitrarily high success probability. On the other hand, we show that for there is necessarily a large subset of GF-non-learnable targets. In particular, the set of learnable targets is not dense in , and any subset of homeomorphic to the -dimensional sphere contains non-learnable targets. Finally, we observe that the model in our main theorem on almost guaranteed two-parameter learning is constructed using a hierarchical procedure and as a result is not expressible by a single elementary function. We show that this limitation is essential in the sense that most models written in terms of elementary functions cannot achieve the learnability demonstrated in this theorem.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3be8b02e-6217-49db-84d0-b1d1d27b1493Builds on3
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Elementary superexpressive activationsDmitry YarotskyICML 2021 · 46 citations
- Constructive Universal High-Dimensional Distribution Generation through Deep ReLU NetworksDmytro Perekrestenko, Stephan Müller, Helmut BölcskeiICML 2020 · 15 citations
Related papers
- Mean-Field Analysis for Learning Subspace-Sparse Polynomials with Gaussian InputZiang Chen, Rong GeNeurIPS 2024 · 1 citation
- The Computational Advantage of Depth in Learning High-Dimensional Hierarchical TargetsYatin Dandi, Luca Pesce, Lenka Zdeborová, Florent KrzakalaNeurIPS 2025 · 3 citations
- Neural Networks Learn Generic Multi-Index Models Near Information-Theoretic LimitBohan Zhang, Zihao Wang, Hengyu Fu, Jason D. LeeICLR 2026 · 3 citations
- A solvable model of learning generative diffusion: theory and insightsHugo Cui, Cengiz Pehlevan, Yue M. LuNeurIPS 2025 · 11 citations
- Deep Network Approximation in Terms of Intrinsic ParametersZuowei Shen, Haizhao Yang, Shijun ZhangICML 2022 · 13 citations
