Mean-Field Langevin Dynamics for Signed Measures via a Bilevel Approach
Guillaume Wang, Alireza Mousavi-Hosseini, Lénaïc Chizat
Abstract
Mean-field Langevin dynamics (MLFD) is a class of interacting particle methods that tackle convex optimization over probability measures on a manifold, which are scalable, versatile, and enjoy computational guarantees. However, some important problems -- such as risk minimization for infinite width two-layer neural networks, or sparse deconvolution -- are originally defined over the set of signed, rather than probability, measures. In this paper, we investigate how to extend the MFLD framework to convex optimization problems over signed measures. Among two known reductions from signed to probability measures -- the lifting and the bilevel approaches -- we show that the bilevel reduction leads to stronger guarantees and faster rates (at the price of a higher per-iteration complexity). In particular, we investigate the convergence rate of MFLD applied to the bilevel reduction in the low-noise regime and obtain two results. First, this dynamics is amenable to an annealing schedule, adapted from Suzuki et al. (2023), that results in improved convergence rates to a fixed multiplicative accuracy. Second, we investigate the problem of learning a single neuron with the bilevel approach and obtain local exponential convergence rates that depend polynomially on the dimension and noise level (to compare with the exponential dependence that would result from prior analyses).
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 84fdeb89-d23c-4240-a5ed-b30704fe11b6Cited by top-tier papers3
- From Information to Generative Exponent: Learning Rate Induces Phase Transitions in SGDKonstantinos C. Tsiolis, Alireza Mousavi-Hosseini, Murat A. ErdogduNeurIPS 2025 · 2 citations
- Learning Multi-Index Models with Neural Networks via Mean-Field Langevin DynamicsAlireza Mousavi-Hosseini, Denny Wu, Murat A. ErdogduICLR 2025
- Robust Feature Learning for Multi-Index Models in High DimensionsAlireza Mousavi-Hosseini, Adel Javanmard, Murat A. ErdogduICLR 2025
Builds on6
- Smooth Bilevel Programming for Sparse RegularizationClarice Poon, Gabriel PeyréNeurIPS 2021 · 23 citations
- Leveraging the two-timescale regime to demonstrate convergence of neural networksPierre Marion, Raphaël BerthierNeurIPS 2023 · 19 citations
- Feature learning via mean-field Langevin dynamics: classifying sparse parities and beyondTaiji Suzuki, Denny Wu, Kazusato Oko, Atsushi NitandaNeurIPS 2023 · 17 citations
- Mean-field Analysis on Two-layer Neural Networks from a Kernel PerspectiveShokichi Takakura, Taiji SuzukiICML 2024 · 12 citations
- Mean-field Langevin dynamics: Time-space discretization, stochastic gradient, and variance reductionTaiji Suzuki, Denny Wu, Atsushi NitandaNeurIPS 2023 · 10 citations
Related papers
- Mirror Mean-Field Langevin DynamicsAnming Gu, Juno KimICML 2026 · 3 citations
- Improved statistical and computational complexity of the mean-field Langevin dynamics under structured dataAtsushi Nitanda, Kazusato Oko, Taiji Suzuki, Denny WuICLR 2024 · 4 citations
- Improved Particle Approximation Error for Mean Field Neural NetworksAtsushi NitandaNeurIPS 2024 · 18 citations
- Particle Dual Averaging: Optimization of Mean Field Neural Network with Global Convergence Rate AnalysisAtsushi Nitanda, Denny Wu, Taiji SuzukiNeurIPS 2021 · 32 citations
- Mean-field Underdamped Langevin Dynamics and its Spacetime DiscretizationQiang Fu, Ashia Camage WilsonICML 2024 · 5 citations
