An Effective Hard Thresholding Method Based on Stochastic Variance Reduction for Nonconvex Sparse Learning
Guannan Liang, Qianqian Tong, Chunjiang Zhu, Jinbo Bi
摘要
We propose a hard thresholding method based on stochastically controlled stochastic gradients (SCSG-HT) to solve a family of sparsity-constrained empirical risk minimization problems. The SCSG-HT uses batch gradients where batch size is pre-determined by the desirable precision tolerance rather than full gradients to reduce the variance in stochastic gradients. It also employs the geometric distribution to determine the number of loops per epoch. We prove that, similar to the latest methods based on stochastic gradient descent or stochastic variance reduction methods, SCSG-HT enjoys a linear convergence rate. However, SCSG-HT now has a strong guarantee to recover the optimal sparse estimator. The computational complexity of SCSG-HT is independent of sample size n when n is larger than 1 , which enhances the scalability to massive-scale problems. Empirical results demonstrate that SCSG-HT outperforms several competitors and decreases the objective value the most with the same computational costs.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper1
相关 Paper
- Variance Reduction With Sparse GradientsMelih Elibol, Lihua Lei, Michael I. JordanICLR 2020 · 被引用 25 次
- New Insight of Variance reduce in Zero-Order Hard-Thresholding: Mitigating Gradient Error and Expansivity ContradictionsXinzhe Yuan, William de Vazelhes, Bin Gu, Huan XiongICLR 2024 · 被引用 1 次
- Almost Tune-Free Variance ReductionBingcong Li, Lingda Wang, Georgios B. GiannakisICML 2020 · 被引用 20 次
- Stochastic Frank-Wolfe for Constrained Finite-Sum MinimizationGeoffrey Négiar, Gideon Dresdner, Alicia Y. Tsai, Laurent El Ghaoui 等ICML 2020 · 被引用 29 次
- Sparse Regression with Constraints for -Mixing Time Series: Algorithms and GuaranteesRuoxin Yuan, Lijun DingICML 2026
