Stochastic Subspace Cubic Newton Method
Filip Hanzely, Nikita Doikov, Yurii E. Nesterov, Peter Richtárik
摘要
In this paper, we propose a new randomized second-order optimization algorithm---Stochastic Subspace Cubic Newton (SSCN)---for minimizing a high dimensional convex function . Our method can be seen both as a stochastic extension of the cubically-regularized Newton method of Nesterov and Polyak (2006), and a second-order enhancement of stochastic subspace descent of Kozak et al. (2019). We prove that as we vary the minibatch size, the global convergence rate of SSCN interpolates between the rate of stochastic coordinate descent (CD) and the rate of cubic regularized Newton, thus giving new insights into the connection between first and second-order methods. Remarkably, the local convergence rate of SSCN matches the rate of stochastic subspace descent applied to the problem of minimizing the quadratic function , where is the minimizer of , and hence depends on the properties of at the optimum only. Our numerical experiments show that SSCN outperforms non-accelerated first-order CD algorithms while being competitive to their accelerated variants.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Distributed Second Order Methods with Fast Rates and Compressed CommunicationRustem Islamov, Xun Qian, Peter RichtárikICML 2021 · 被引用 56 次
- Finding Second-Order Stationary Points in Nonconvex-Strongly-Concave Minimax OptimizationLuo Luo, Yujun Li, Cheng ChenNeurIPS 2022 · 被引用 22 次
- Spectral Preconditioning for Gradient Methods on Graded Non-convex FunctionsNikita Doikov, Sebastian U. Stich, Martin JaggiICML 2024 · 被引用 10 次
- Convex optimization based on global lower second-order modelsNikita Doikov, Yurii E. NesterovNeurIPS 2020 · 被引用 9 次
- Turbocharging Gaussian Process Inference with Approximate Sketch-and-ProjectPratik Rathore, Zachary Frangella, Sachin Garg, Shaghayegh Fazliani 等NeurIPS 2025 · 被引用 8 次
相关 Paper
- A Damped Newton Method Achieves Global and Local Quadratic Convergence RateSlavomír Hanzely, Dmitry Kamzolov, Dmitry Pasechnyuk, Alexander V. Gasnikov 等NeurIPS 2022 · 被引用 1 次
- Escaping Saddle-Point Faster under Interpolation-like ConditionsAbhishek Roy, Krishnakumar Balasubramanian, Saeed Ghadimi, Prasant MohapatraNeurIPS 2020 · 被引用 8 次
- Faster Differentially Private Convex Optimization via Second-Order MethodsArun Ganesh, Mahdi Haghifam, Thomas Steinke, Abhradeep Guha ThakurtaNeurIPS 2023 · 被引用 18 次
- Improving Convergence Guarantees of Random Subspace Second-order Algorithm for Nonconvex OptimizationRei Higuchi, Pierre-Louis Poirion, Akiko TakedaICLR 2025
- OPTAMI: Global Superlinear Convergence of High-order MethodsDmitry Kamzolov, Artem Agafonov, Dmitry Pasechnyuk, Alexander V. Gasnikov 等ICLR 2025
