SC2023Top-tier venue
Optimizing High-Performance Linpack for Exascale Accelerated Architectures
Noel Chalmers, Jakub Kurzak, Damon McDougall, Paul T. Bauman
Abstract
We detail the performance optimizations made in rocHPL, AMD's open-source implementation of the High-Performance Linpack (HPL) benchmark targeting accelerated node architectures designed for exascale systems such as the Frontier supercomputer. The implementation leverages the high-throughput GPU accelerators on the node via highly optimized linear algebra libraries, as well as the entire CPU socket to perform latency-sensitive factorization phases. We detail novel performance improvements such as a multithreaded approach to computing the panel factorization phase on the CPU, time-sharing of CPU cores between processes on the node, as well as several optimizations which hide MPI communication. We present some performance results of this implementation of the HPL benchmark on a single node of the Frontier early access cluster at Oak Ridge National Laboratory, as well as scaling to multiple nodes.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 96720bb7-504f-4cd5-afe3-3697d57e5451Cited by top-tier papers2
- KAMI: Communication-Avoiding General Matrix Multiplication within a Single GPUHemeng Wang, Yang Du, Sidu Li, Xiaowen Tian et al.SC 2025 · 4 citations
- Trojan Horse: Aggregate-and-Batch for Scaling Up Sparse Direct Solvers on GPU ClustersYida Li, Siwei Zhang, Yiduo Niu, Yang Du et al.PPoPP 2026 · 1 citation
Related papers
- Insights from Optimizing HPL Performance on Exascale Systems: A Comparative Analysis of Panel FactorizationHao Lu, Michael A. Matheson, Noel Chalmers, Aditya Kashi et al.SC 2025 · 1 citation
- Climbing the Summit and Pushing the Frontier of Mixed Precision Benchmarks at Extreme ScaleHao Lu, Michael A. Matheson, Vladyslav Oles, J. Austin Ellis et al.SC 2022 · 8 citations
- Frontier: Exploring ExascaleScott Atchley, Christopher Zimmer, John Lange, David E. Bernholdt et al.SC 2023 · 91 citations
- A Digital Twin Framework for Liquid-cooled Supercomputers as Demonstrated at ExascaleWesley Brewer, Matthias Maiterth, Vineet Kumar, Rafal P. Wojda et al.SC 2024 · 22 citations
- 5 ExaFlop/s HPL-MxP Benchmark with Linear Scalability on the 40-Million-Core Sunway SupercomputerRongfen Lin, Xinhui Yuan, Wei Xue, Wanwang Yin et al.SC 2023 · 11 citations
