Cogradient Descent for Bilinear Optimization
Li'an Zhuo, Baochang Zhang, Linlin Yang, Hanlin Chen, Qixiang Ye, David S. Doermann, Rongrong Ji, Guodong Guo
摘要
Conventional learning methods simplify the bilinear model by regarding two intrinsically coupled factors independently, which degrades the optimization procedure. One reason lies in the insufficient training due to the asynchronous gradient descent, which results in vanishing gradients for the coupled variables. In this paper, we introduce a Cogradient Descent algorithm (CoGD) to address the bilinear problem, based on a theoretical framework to coordinate the gradient of hidden variables via a projection function. We solve one variable by considering its coupling relationship with the other, leading to a synchronous gradient descent to facilitate the optimization procedure. Our algorithm is applied to solve problems with one variable under the sparsity constraint, which is widely used in the learning paradigm. We validate our CoGD considering an extensive set of applications including image reconstruction, inpainting, and network pruning. Experiments show that it improves the state-of-the-art by a significant margin 1 .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- SCOP: Scientific Control for Reliable Neural Network PruningYehui Tang, Yunhe Wang, Yixing Xu, Dacheng Tao 等NeurIPS 2020 · 被引用 208 次
- Searching for Low-Bit Weights in Quantized Neural NetworksZhaohui Yang, Yunhe Wang, Kai Han, Chunjing Xu 等NeurIPS 2020 · 被引用 103 次
- GradAug: A New Regularization Method for Deep Neural NetworksTaojiannan Yang, Sijie Zhu, Chen ChenNeurIPS 2020 · 被引用 43 次
- IDARTS: Interactive Differentiable Architecture SearchSong Xue, Runqi Wang, Baochang Zhang, Tian Wang 等ICCV 2021 · 被引用 10 次
- Layer-Wise Searching for 1-Bit DetectorsSheng Xu, Junhe Zhao, Jinhu Lü, Baochang Zhang 等CVPR 2021
它引用的顶会 Paper3
- Noise Flow: Noise Modeling With Conditional Normalizing FlowsAbdelrahman Abdelhamed, Marcus A. Brubaker, Michael S. BrownICCV 2019 · 被引用 199 次
- Accelerate CNN via Recursive Bayesian PruningYuefu Zhou, Ya Zhang, Yanfeng Wang, Qi TianICCV 2019 · 被引用 64 次
- Solving Vision Problems via FilteringSean I. Young, Aous Thabit Naman, Bernd Girod, David TaubmanICCV 2019 · 被引用 4 次
相关 Paper
- Granger Components Analysis: Unsupervised learning of latent temporal dependenciesJacek DmochowskiNeurIPS 2023 · 被引用 2 次
- GAN-Based Projector for Faster Recovery With Convergence Guarantees in Linear Inverse ProblemsAnkit Raj, Yuqi Li, Yoram BreslerICCV 2019 · 被引用 61 次
- Nesterov Meets Optimism: Rate-Optimal Separable Minimax OptimizationChris Junchi Li, Huizhuo Yuan, Gauthier Gidel, Quanquan Gu 等ICML 2023 · 被引用 8 次
- GDA-AM: On the Effectiveness of Solving Min-Imax Optimization via Anderson MixingHuan He, Shifan Zhao, Yuanzhe Xi, Joyce C. Ho 等ICLR 2022 · 被引用 12 次
- Smoothing Proximal Gradient Methods for Nonsmooth Sparsity Constrained Optimization: Optimality Conditions and Global ConvergenceGanzhao YuanICML 2024 · 被引用 6 次
