Two-layer neural network on infinite dimensional data: global optimization guarantee in the mean-field regime
Naoki Nishikawa, Taiji Suzuki, Atsushi Nitanda, Denny Wu
Abstract
The analysis of neural network optimization in the mean-field regime is important as the setting allows for feature learning. The existing theory has been developed mainly for neural networks in finite dimensions, i.e. each neuron has a finite-dimensional parameter. However, the setting of infinite-dimensional input naturally arises in machine learning problems such as nonparametric functional data analysis and graph classification. In this paper, we develop a new mean-field analysis of a two-layer neural network in an infinite-dimensional parameter space. We first give a generalization error bound, which shows that the regularized empirical risk minimizer properly generalizes when the data size is sufficiently large, despite the neurons being infinite-dimensional. Next, we present two gradient-based optimization algorithms for infinite-dimensional mean-field networks, by extending the recently developed particle optimization framework to the infinite-dimensional setting. We show that the proposed algorithms converge to the (regularized) global optimal solution, and moreover, their rates of convergence are of polynomial order in the online setting and exponential order in the finite sample setting, respectively. To the best of our knowledge, this is the first quantitative global optimization guarantee of a neural network on infinite-dimensional input and in the presence of feature learning.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3f4c78f2-4c09-4a64-b205-1469d26e502dCited by top-tier papers6
- Rethinking Information-theoretic Generalization: Loss Entropy Induced PAC BoundsYuxin Dong, Tieliang Gong, Hong Chen, Shujian Yu et al.ICLR 2024 · 8 citations
- Primal and Dual Analysis of Entropic Fictitious Play for Finite-sum ProblemsAtsushi Nitanda, Kazusato Oko, Denny Wu, Nobuhito Takenouchi et al.ICML 2023 · 4 citations
- Generalization Error of Graph Neural Networks in the Mean-field RegimeGholamali Aminian, Yixuan He, Gesine Reinert, Lukasz Szpruch et al.ICML 2024 · 4 citations
- Transformers as Measure-Theoretic Associative Memory: A Statistical Perspective and Minimax OptimalityRyotaro Kawata, Taiji SuzukiICLR 2026 · 3 citations
- Which Algorithms Have Tight Generalization Bounds?Michael Gastpar, Ido Nachum, Jonathan Shafer, Thomas WeinbergerNeurIPS 2025 · 1 citation
Builds on3
- Tensor Programs IV: Feature Learning in Infinite-Width Neural NetworksGreg Yang, Edward J. HuICML 2021 · 242 citations
- Deep Learning for Functional Data Analysis with Adaptive Basis LayersJunwen Yao, Jonas Mueller, Jane-Ling WangICML 2021 · 40 citations
- Particle Stochastic Dual Coordinate Ascent: Exponential convergent algorithm for mean field neural network optimizationKazusato Oko, Taiji Suzuki, Atsushi Nitanda, Denny WuICLR 2022 · 8 citations
Related papers
- Particle Dual Averaging: Optimization of Mean Field Neural Network with Global Convergence Rate AnalysisAtsushi Nitanda, Denny Wu, Taiji SuzukiNeurIPS 2021 · 32 citations
- Global Convergence of Three-layer Neural Networks in the Mean Field RegimeHuy Tuan Pham, Phan-Minh NguyenICLR 2021 · 23 citations
- Deep Neural Network Regression with Functional CovariatesHang Zhou, Ju-Sheng Hong, Xiucai Ding, Jane-Ling WangICML 2026
- On feature learning in neural networks with global convergence guaranteesZhengdao Chen, Eric Vanden-Eijnden, Joan BrunaICLR 2022 · 15 citations
- Generalization bound of globally optimal non-convex neural network training: Transportation map estimation by infinite dimensional Langevin dynamicsTaiji SuzukiNeurIPS 2020 · 25 citations
