Learning Sparse Approximate Inverse Preconditioners for Conjugate Gradient Solvers on GPUs
Zhehao Li, Kangbo Lyu, Yixuan Li, Tao Du, Ligang Liu
摘要
The conjugate gradient solver (CG) is a prevalent method for solving symmetric and positive definite linear systems Ax=b, where effective preconditioners are crucial for fast convergence. Traditional preconditioners rely on prescribed algorithms to offer rigorous theoretical guarantees, while limiting their ability to exploit optimization from data. Existing learning-based methods often utilize Graph Neural Networks (GNNs) to improve the performance and speed up the construction. However, their reliance on incomplete factorization leads to significant challenges: the associated triangular solve hinders GPU parallelization in practice, and introduces long-range dependencies which are difficult for GNNs to model. To address these issues, we propose a learning-based method to generate GPU-friendly preconditioners, particularly using GNNs to construct Sparse Approximate Inverse (SPAI) preconditioners, which avoids triangular solves and requires only two matrix-vector products at each CG step. The locality of matrix-vector product is compatible with the local propagation mechanism of GNNs. The flexibility of GNNs also allows our approach to be applied in a wide range of scenarios. Furthermore, we introduce a statistics-based scale-invariant loss function. Its design matches CG's property that the convergence rate depends on the condition number, rather than the absolute scale of A, leading to improved performance of the learned preconditioner. Evaluations on three PDE-derived datasets and one synthetic dataset demonstrate that our method outperforms standard preconditioners (Diagonal, IC, and traditional SPAI) and previous learning-based preconditioners on GPUs. We reduce solution time on GPUs by 40%-53% (68%-113% faster), along with better condition numbers and superior generalization performance. Source code available at https://github.com/Adversarr/LearningSparsePreconditioner4GPU
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Mixture-of-Experts Operator Transformer for Large-Scale PDE Pre-TrainingHong Wang, Haiyang Xin, Jie Wang, Xuanze Yang 等NeurIPS 2025 · 被引用 15 次
- SymMaP: Improving Computational Efficiency in Linear Solvers through Symbolic PreconditioningHong Wang, Jie Wang, Minghao Ma, Haoran Shao 等NeurIPS 2025 · 被引用 6 次
- STNet: Spectral Transformation Network for Solving Operator Eigenvalue ProblemHong Wang, Yixuan Jiang, Jie Wang, Xinyi Li 等NeurIPS 2025 · 被引用 4 次
它引用的顶会 Paper6
- Optimization-Based Algebraic Multigrid Coarsening Using Reinforcement LearningAli Taghibakhshi, Scott P. MacLachlan, Luke N. Olson, Matthew WestNeurIPS 2021 · 被引用 43 次
- Learning Preconditioners for Conjugate Gradient PDE SolversYichen Li, Peter Yichen Chen, Tao Du, Wojciech MatusikICML 2023 · 被引用 38 次
- Neural Krylov Iteration for Accelerating Linear System SolvingJian Luo, Jie Wang, Hong Wang, Huanshuo Dong 等NeurIPS 2024 · 被引用 23 次
- Neural operators meet conjugate gradients: The FCG-NO method for efficient PDE solvingAlexander Rudikov, Vladimir Fanaskov, Ekaterina A. Muravleva, Yuri M. Laevsky 等ICML 2024 · 被引用 14 次
- A Neural-Preconditioned Poisson Solver for Mixed Dirichlet and Neumann Boundary ConditionsKai Weixian Lan, Elias Gueidon, Ayano Kaneda, Julian Panetta 等ICML 2024 · 被引用 4 次
相关 Paper
- Graph Neural Preconditioners for Iterative Solutions of Sparse Linear SystemsJie ChenICLR 2025
- Extending Sparse Patterns to Improve Inverse Preconditioning on GPU ArchitecturesSergi Laut, Ricard Borrell, Marc CasasHPDC 2024 · 被引用 3 次
- Sparsified Preconditioned Conjugate Gradient Solver on GPUsDa Ma, Khalid Ahmad, Kazem Cheshmi, Hari Sundar 等SC 2025 · 被引用 1 次
- Learning Algebraic Multigrid Using Graph Neural NetworksIlay Luz, Meirav Galun, Haggai Maron, Ronen Basri 等ICML 2020 · 被引用 95 次
- RAPNet: Accelerating Algebraic Multigrid with Learned Sparse CorrectionsYali Fink, Ido Ben-Yair, Lars Ruthotto, Eran TreisterICML 2026
