SC2025Top-tier venue
Utilizing Sparsity in the GPU-accelerated Assembly of Schur Complement Matrices in Domain Decomposition Methods
Jakub Homola, Ondrej Meca, Lubomír Ríha, Tomás Brzobohatý
Abstract
Schur complement matrices emerge in many domain decomposition methods that can solve complex engineering problems using supercomputers. Today, as most of the high-performance clusters' performance lies in GPUs, these methods should also be accelerated.
Typically, the offloaded components are the explicitly assembled dense Schur complement matrices used later in the iterative solver for multiplication with a vector. As the explicit assembly is expensive, it represents a significant overhead associated with this approach to acceleration. It has already been shown that the overhead can be minimized by assembling the Schur complements directly on the GPU.
This paper shows that the GPU assembly can be further improved by wisely utilizing the sparsity of the input matrices. In the context of FETI methods, we achieved a speedup of 5.1 in the GPU section of the code and 3.3 for the whole assembly, making the acceleration beneficial from as few as 10 iterations.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Related papers
- Solving Linear Systems on a GPU with Hierarchically Off-Diagonal Low-Rank ApproximationsChao Chen, Per-Gunnar MartinssonSC 2022 · 4 citations
- Second-order Stencil Descent for Interior-point HyperelasticityLei Lan, Minchen Li, Chenfanfu Jiang, Huamin Wang et al.SIGGRAPH 2023 · 33 citations
- Addressing Irregular Patterns of Matrix Computations on GPUs and Their Impact on Applications Powered by Sparse Direct SolversAhmad Abdelfattah, Pieter Ghysels, Wajih Boukaram, Stanimire Tomov et al.SC 2022 · 4 citations
- A GPU-based multilevel additive schwarz preconditioner for cloth and deformable body simulationBotao Wu, Zhendong Wang, Huamin WangSIGGRAPH 2022 · 48 citations
- Extending Sparse Patterns to Improve Inverse Preconditioning on GPU ArchitecturesSergi Laut, Ricard Borrell, Marc CasasHPDC 2024 · 3 citations
