Extending Sparse Patterns to Improve Inverse Preconditioning on GPU Architectures
Sergi Laut, Ricard Borrell, Marc Casas
摘要
Graphic Processing Units (GPUs) have become a key component of high-end computing infrastructures due to their massively parallel architecture, which delivers large floating-point operations per cycle rates. Many scientific workloads benefit from GPUs and, in particular, numerical methods solving linear systems of equations Ax = b typically run on GPUs. Among them, the Conjugate Gradient (CG) method, which targets linear systems with Symmetric and Positive Definite (SPD) matrices, runs on GPUs using its preconditioned form. However, state-of-the-art preconditioning techniques like the Factorized Sparse Approximate Inverse (FSAI) preconditioner ignore the benefits of data coalescence and locality on GPU architectures and leave substantial performance on the table. These approaches are exclusively based on numerical criteria.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper1
问问它们各自怎么用它相关 Paper
- Cache-aware Sparse Patterns for the Factorized Sparse Approximate Inverse PreconditionerSergi Laut, Ricard Borrell, Marc CasasHPDC 2021 · 被引用 3 次
- Communication-aware Sparse Patterns for the Factorized Approximate Inverse PreconditionerSergi Laut, Marc Casas, Ricard BorrellHPDC 2022 · 被引用 4 次
- Learning Sparse Approximate Inverse Preconditioners for Conjugate Gradient Solvers on GPUsZhehao Li, Kangbo Lyu, Yixuan Li, Tao Du 等NeurIPS 2025 · 被引用 5 次
- Sparsified Preconditioned Conjugate Gradient Solver on GPUsDa Ma, Khalid Ahmad, Kazem Cheshmi, Hari Sundar 等SC 2025 · 被引用 1 次
- CPU- and GPU-initiated Communication Strategies for Conjugate Gradient Methods on Large GPU ClustersJames D. Trotter, Sinan Ekmekçibasi, Dogan Sagbili, Johannes Langguth 等SC 2025 · 被引用 2 次
