RandProx: Primal-Dual Optimization Algorithms with Randomized Proximal Updates
Laurent Condat, Peter Richtárik
摘要
Proximal splitting algorithms are well suited to solving large-scale nonsmooth optimization problems, in particular those arising in machine learning. We propose a new primal-dual algorithm, in which the dual update is randomized; equivalently, the proximity operator of one of the function in the problem is replaced by a stochastic oracle. For instance, some randomly chosen dual variables, instead of all, are updated at each iteration. Or, the proximity operator of a function is called with some small probability only. A nonsmooth variance-reduction technique is implemented so that the algorithm finds an exact minimizer of the general problem involving smooth and nonsmooth functions, possibly composed with linear operators. We derive linear convergence results in presence of strong convexity; these results are new even in the deterministic case, when our algorithms reverts to the recently proposed Primal-Dual Davis-Yin algorithm. Some randomized algorithms of the literature are also recovered as particular cases (e.g., Point-SAGA). But our randomization technique is general and encompasses many unbiased mechanisms beyond sampling and probabilistic updates, including compression. Since the convergence speed depends on the slowest among the primal and dual contraction mechanisms, the iteration complexity might remain the same when randomness is used. On the other hand, the computation complexity can be significantly reduced. Overall, randomness helps getting faster algorithms. This has long been known for stochastic-gradient-type algorithms, and our work shows that this fully applies in the more general primal-dual setting as well. 1 2 x ′ -x 2 . This operator has a closed form for many functions of practical interest (Parikh & Boyd, 2014; Pustelnik & Condat, 2017; Gheche et al., 2018) , see also the website http://proximity-operator.net . In addition, the Moreau identity holds: where φ * : x ∈ X → sup x ′ ∈X x, x ′ -φ(x ′ ) denotes the conjugate function of φ (Bauschke & Combettes, 2017). Thus, one can compute the proximity operator of φ from the one of φ * , and conversely.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- Federated Optimization with Doubly Regularized Drift CorrectionXiaowen Jiang, Anton Rodomanov, Sebastian U. StichICML 2024 · 被引用 18 次
- Stabilized Proximal-Point Methods for Federated OptimizationXiaowen Jiang, Anton Rodomanov, Sebastian U. StichNeurIPS 2024 · 被引用 13 次
- Tighter Analysis for ProxSkipZhengmian Hu, Heng HuangICML 2023 · 被引用 11 次
- SCAFFLSA: Taming Heterogeneity in Federated Linear Stochastic Approximation and TD LearningPaul Mangold, Sergey Samsonov, Safwan Labbi, Ilya Levin 等NeurIPS 2024 · 被引用 10 次
- Multiplayer Federated Learning: Reaching Equilibrium with Less CommunicationTaeHo Yoon, Sayantan Choudhury, Nicolas LoizouNeurIPS 2025 · 被引用 7 次
它引用的顶会 Paper9
- SCAFFOLD: Stochastic Controlled Averaging for Federated LearningSai Praneeth Karimireddy, Satyen Kale, Mehryar Mohri, Sashank J. Reddi 等ICML 2020 · 被引用 3,875 次
- Minibatch vs Local SGD for Heterogeneous Distributed LearningBlake E. Woodworth, Kumar Kshitij Patel, Nati SrebroNeurIPS 2020 · 被引用 231 次
- Lower Bounds and Optimal Algorithms for Personalized Federated LearningFilip Hanzely, Slavomír Hanzely, Samuel Horváth, Peter RichtárikNeurIPS 2020 · 被引用 215 次
- ProxSkip: Yes! Local Gradient Steps Provably Lead to Communication Acceleration! Finally!Konstantin Mishchenko, Grigory Malinovsky, Sebastian U. Stich, Peter RichtárikICML 2022 · 被引用 200 次
- From Local SGD to Local Fixed-Point Methods for Federated LearningGrigory Malinovskiy, Dmitry Kovalev, Elnur Gasanov, Laurent Condat 等ICML 2020 · 被引用 135 次
相关 Paper
- Variance Reduction via Primal-Dual Accelerated Dual Averaging for Nonsmooth Convex Finite-SumsChaobing Song, Stephen J. Wright, Jelena DiakonikolasICML 2021 · 被引用 22 次
- Stochastic Smoothed Primal-Dual Algorithms for Nonconvex Optimization with Linear Inequality ConstraintsRuichuan Huang, Jiawei Zhang, Ahmet AlacaogluICML 2025
- Hybrid Variance-Reduced SGD Algorithms For Minimax Problems with Nonconvex-Linear FunctionQuoc Tran-Dinh, Deyi Liu, Lam M. NguyenNeurIPS 2020 · 被引用 28 次
- S-D-RSM: Stochastic Distributed Regularized Splitting Method for Large-Scale Convex Optimization ProblemsMaoran Wang, Xingju Cai, Yongxin ChenAAAI 2026
- Kill a Bird with Two Stones: Closing the Convergence Gaps in Non-Strongly Convex Optimization by Directly Accelerated SVRG with Double Compensation and SnapshotsYuanyuan Liu, Fanhua Shang, Weixin An, Hongying Liu 等ICML 2022 · 被引用 2 次
