Cache-aware Sparse Patterns for the Factorized Sparse Approximate Inverse Preconditioner
Sergi Laut, Ricard Borrell, Marc Casas
Abstract
Conjugate Gradient is a widely used iterative method to solve linear systems 𝐴𝑥 = 𝑏 with matrix 𝐴 being symmetric and positive definite. Part of its effectiveness relies on finding a suitable preconditioner that accelerates its convergence. Factorized Sparse Approximate Inverse (FSAI) preconditioners are a prominent and easily parallelizable option. An essential element of a FSAI preconditioner is the definition of its sparse pattern, which constraints the approximation of the inverse 𝐴 -1 . This definition is generally based on numerical criteria. In this paper we introduce complementary architecture-aware criteria to increase the numerical effectiveness of the preconditioner without incurring in significant performance costs. In particular, we define cache-aware pattern extensions that do not trigger additional cache misses when accessing vector 𝑥 in the 𝑦 = 𝐴𝑥 Sparse Matrix-Vector (SpMV) kernel. As a result, we obtain very significant reductions in terms of average solution time ranging between 12.94% and 22.85% on three different architectures -Intel Skylake, POWER9 and A64FX -over a set of 72 test matrices.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 638716ad-e520-44ae-a0b0-89009acf6613Related papers
- Communication-aware Sparse Patterns for the Factorized Approximate Inverse PreconditionerSergi Laut, Marc Casas, Ricard BorrellHPDC 2022 · 4 citations
- Extending Sparse Patterns to Improve Inverse Preconditioning on GPU ArchitecturesSergi Laut, Ricard Borrell, Marc CasasHPDC 2024 · 3 citations
- Learning Sparse Approximate Inverse Preconditioners for Conjugate Gradient Solvers on GPUsZhehao Li, Kangbo Lyu, Yixuan Li, Tao Du et al.NeurIPS 2025 · 5 citations
- Me-MPK: Accelerating Krylov Subspace Solvers via Memory-efficient Matrix-Power KernelHaozhong Qiu, Chuanfu Xu, Jianbin Fang, Shengguo Li et al.DAC 2025 · 1 citation
- Sparsified Preconditioned Conjugate Gradient Solver on GPUsDa Ma, Khalid Ahmad, Kazem Cheshmi, Hari Sundar et al.SC 2025 · 1 citation
