SC2022Top-tier venue
Solving Linear Systems on a GPU with Hierarchically Off-Diagonal Low-Rank Approximations
Chao Chen, Per-Gunnar Martinsson
Abstract
We are interested in solving linear systems arising from three applications: (1) kernel methods in machine learning, (2) discretization of boundary integral equations from mathematical physics, and (3) Schur complements formed in the factorization of many large sparse matrices. The coefficient matrices are often data-sparse in the sense that their off-diagonal blocks have low numerical ranks; specifically, we focus on “hierarchically off-diagonal low-rank (HODLR)” matrices. We introduce algorithms for factorizing HODLR matrices and for applying the factorizations on a GPU. The algorithms leverage the efficiency of batched dense linear algebra, and they scale nearly linearly with the matrix size when the numerical ranks are fixed. The accuracy of the HODLR-matrix approximation is a tunable parameter, so we can construct high-accuracy fast direct solvers or low-accuracy robust preconditioners. Numerical results show that we can solve problems with several millions of unknowns in a couple of seconds on a single GPU.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a6b3d394-63d2-4234-9107-571e4c69b62cBuilds on1
Related papers
- Scalable Linear Time Dense Direct Solver for 3-D Problems without Trailing Sub-Matrix DependenciesQianxiang Ma, Sameer Deshmukh, Rio YokotaSC 2022 · 7 citations
- Addressing Irregular Patterns of Matrix Computations on GPUs and Their Impact on Applications Powered by Sparse Direct SolversAhmad Abdelfattah, Pieter Ghysels, Wajih Boukaram, Stanimire Tomov et al.SC 2022 · 4 citations
- Kernel Methods Through the Roof: Handling Billions of Points EfficientlyGiacomo Meanti, Luigi Carratino, Lorenzo Rosasco, Alessandro RudiNeurIPS 2020 · 138 citations
- Utilizing Sparsity in the GPU-accelerated Assembly of Schur Complement Matrices in Domain Decomposition MethodsJakub Homola, Ondrej Meca, Lubomír Ríha, Tomás BrzobohatýSC 2025 · 1 citation
- Lightning-fast Boundary Element MethodJiong Chen, Florian Schäfer, Mathieu DesbrunSIGGRAPH 2025 · 3 citations
