Stochastic Approximation Algorithms for Systems of Interacting Particles
Mohammad Reza Karimi Jaghargh, Ya-Ping Hsieh, Andreas Krause
摘要
Interacting particle systems have proven highly successful in various machine learning tasks, including approximate Bayesian inference and neural network optimization. However, the analysis of these systems often relies on the simplifying assumption of the mean-field limit, where particle numbers approach infinity and infinitesimal step sizes are used. In practice, discrete time steps, finite particle numbers, and complex integration schemes are employed, creating a theoretical gap between continuous-time and discrete-time processes. In this paper, we present a novel framework that establishes a precise connection between these discretetime schemes and their corresponding mean-field limits in terms of convergence properties and asymptotic behavior. By adopting a dynamical system perspective, our framework seamlessly integrates various numerical schemes that are typically analyzed independently. For example, our framework provides a unified treatment of optimizing an infinite-width two-layer neural network and sampling via Stein Variational Gradient descent, which were previously studied in isolation.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper9
- A Non-Asymptotic Analysis for Stein Variational Gradient DescentAnna Korba, Adil Salim, Michael Arbel, Giulia Luise 等NeurIPS 2020 · 被引用 102 次
- The Limits of Min-Max Optimization Algorithms: Convergence to Spurious Non-Critical SetsYa-Ping Hsieh, Panayotis Mertikopoulos, Volkan CevherICML 2021 · 被引用 96 次
- SVGD as a kernelized Wasserstein gradient flow of the chi-squared divergenceSinho Chewi, Thibaut Le Gouic, Chen Lu, Tyler Maunu 等NeurIPS 2020 · 被引用 92 次
- A mean-field analysis of two-player zero-sum gamesCarles Domingo-Enrich, Samy Jelassi, Arthur Mensch, Grant M. Rotskoff 等NeurIPS 2020 · 被引用 56 次
- A Finite-Particle Convergence Rate for Stein Variational Gradient DescentJiaxin Shi, Lester MackeyNeurIPS 2023 · 被引用 34 次
相关 Paper
- Quantitative Propagation of Chaos for SGD in Wide Neural NetworksValentin De Bortoli, Alain Durmus, Xavier Fontaine, Umut SimsekliNeurIPS 2020 · 被引用 36 次
- A Dynamical Central Limit Theorem for Shallow Neural NetworksZhengdao Chen, Grant M. Rotskoff, Joan Bruna, Eric Vanden-EijndenNeurIPS 2020 · 被引用 33 次
- Towards a General Theory of Infinite-Width Limits of Neural ClassifiersEugene A. GolikovICML 2020 · 被引用 10 次
- Improved statistical and computational complexity of the mean-field Langevin dynamics under structured dataAtsushi Nitanda, Kazusato Oko, Taiji Suzuki, Denny WuICLR 2024 · 被引用 4 次
- Mean-field Langevin dynamics: Time-space discretization, stochastic gradient, and variance reductionTaiji Suzuki, Denny Wu, Atsushi NitandaNeurIPS 2023 · 被引用 10 次
