Two-layer neural network on infinite dimensional data: global optimization guarantee in the mean-field regime
Naoki Nishikawa, Taiji Suzuki, Atsushi Nitanda, Denny Wu
摘要
The analysis of neural network optimization in the mean-field regime is important as the setting allows for feature learning. The existing theory has been developed mainly for neural networks in finite dimensions, i.e. each neuron has a finite-dimensional parameter. However, the setting of infinite-dimensional input naturally arises in machine learning problems such as nonparametric functional data analysis and graph classification. In this paper, we develop a new mean-field analysis of a two-layer neural network in an infinite-dimensional parameter space. We first give a generalization error bound, which shows that the regularized empirical risk minimizer properly generalizes when the data size is sufficiently large, despite the neurons being infinite-dimensional. Next, we present two gradient-based optimization algorithms for infinite-dimensional mean-field networks, by extending the recently developed particle optimization framework to the infinite-dimensional setting. We show that the proposed algorithms converge to the (regularized) global optimal solution, and moreover, their rates of convergence are of polynomial order in the online setting and exponential order in the finite sample setting, respectively. To the best of our knowledge, this is the first quantitative global optimization guarantee of a neural network on infinite-dimensional input and in the presence of feature learning.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Rethinking Information-theoretic Generalization: Loss Entropy Induced PAC BoundsYuxin Dong, Tieliang Gong, Hong Chen, Shujian Yu 等ICLR 2024 · 被引用 8 次
- Primal and Dual Analysis of Entropic Fictitious Play for Finite-sum ProblemsAtsushi Nitanda, Kazusato Oko, Denny Wu, Nobuhito Takenouchi 等ICML 2023 · 被引用 4 次
- Generalization Error of Graph Neural Networks in the Mean-field RegimeGholamali Aminian, Yixuan He, Gesine Reinert, Lukasz Szpruch 等ICML 2024 · 被引用 4 次
- Transformers as Measure-Theoretic Associative Memory: A Statistical Perspective and Minimax OptimalityRyotaro Kawata, Taiji SuzukiICLR 2026 · 被引用 3 次
- Which Algorithms Have Tight Generalization Bounds?Michael Gastpar, Ido Nachum, Jonathan Shafer, Thomas WeinbergerNeurIPS 2025 · 被引用 1 次
它引用的顶会 Paper3
- Tensor Programs IV: Feature Learning in Infinite-Width Neural NetworksGreg Yang, Edward J. HuICML 2021 · 被引用 242 次
- Deep Learning for Functional Data Analysis with Adaptive Basis LayersJunwen Yao, Jonas Mueller, Jane-Ling WangICML 2021 · 被引用 40 次
- Particle Stochastic Dual Coordinate Ascent: Exponential convergent algorithm for mean field neural network optimizationKazusato Oko, Taiji Suzuki, Atsushi Nitanda, Denny WuICLR 2022 · 被引用 8 次
相关 Paper
- Particle Dual Averaging: Optimization of Mean Field Neural Network with Global Convergence Rate AnalysisAtsushi Nitanda, Denny Wu, Taiji SuzukiNeurIPS 2021 · 被引用 32 次
- Global Convergence of Three-layer Neural Networks in the Mean Field RegimeHuy Tuan Pham, Phan-Minh NguyenICLR 2021 · 被引用 23 次
- Deep Neural Network Regression with Functional CovariatesHang Zhou, Ju-Sheng Hong, Xiucai Ding, Jane-Ling WangICML 2026
- On feature learning in neural networks with global convergence guaranteesZhengdao Chen, Eric Vanden-Eijnden, Joan BrunaICLR 2022 · 被引用 15 次
- Generalization bound of globally optimal non-convex neural network training: Transportation map estimation by infinite dimensional Langevin dynamicsTaiji SuzukiNeurIPS 2020 · 被引用 25 次
