Non-recurring engineering (NRE) best practices: a case study with the NERSC/NVIDIA OpenMP contract
Christopher S. Daley, Annemarie Southwell, Rahulkumar Gayatri, Scott Biersdorfff, Craig Toepfer, Güray Özen, Nicholas J. Wright
摘要
The NERSC supercomputer, Perlmutter, consists of AMD CPUs and NVIDIA GPUs. NERSC users expect to be able to use OpenMP to take advantage of the highly capable GPUs. This paper describes how NERSC/NVIDIA constructed a Non-Recurring Engineering (NRE) contract to add OpenMP GPU-offload support to the NVIDIA HPC compilers. The paper describes how the contract incorporated the strengths of both parties and encouraged collaboration to improve the quality of the final deliverable. We include our best practices and how this particular contract took into account emerging OpenMP specifications, NERSC workload requirements, and how to use OpenMP most efficiently on GPU hardware. This paper includes OpenMP application performance results obtained with the NVIDIA compilers distributed in the NVIDIA HPC SDK.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- ODOS-MPI: HPC-Friendly SmartNIC Offloading of Computation/Communication KernelsMuhammad Usman, Mariano Benito, Sergio Iserte, Antonio J. PeñaSC 2025 · 被引用 5 次
- Static Generation of Efficient OpenMP Offload Data MappingsLuke Marzen, Akash Dutta, Ali JannesariSC 2024 · 被引用 4 次
- Unified Communication Optimization Strategies for Sparse Triangular Solver on CPU and GPU ClustersYang Liu, Nan Ding, Piyush Sao, Samuel Williams 等SC 2023 · 被引用 8 次
- Paths to OpenMP in the kernelJiacheng Ma, Wenyi Wang, Aaron Nelson, Michael Cuevas 等SC 2021 · 被引用 4 次
- Not All GPUs Are Created Equal: Characterizing Variability in Large-Scale, Accelerator-Rich SystemsPrasoon Sinha, Akhil Guliani, Rutwik Jain, Brandon Tran 等SC 2022 · 被引用 31 次
