HPAC-Offload: Accelerating HPC Applications with Portable Approximate Computing on the GPU
Zane Fink, Konstantinos Parasyris, Giorgis Georgakoudis, Harshitha Menon
摘要
The end of Dennard scaling and the slowdown of Moore's law led to a shift in technology trends towards parallel architectures, particularly in HPC systems. To continue providing performance benefits, HPC should embrace Approximate Computing (AC), which trades application quality loss for improved performance. However, existing AC techniques have not been extensively applied and evaluated in state-of-the-art hardware architectures such as GPUs, the primary execution vehicle for HPC applications today.
This paper presents HPAC-Offload, a pragma-based programming model that extends OpenMP offload applications to support AC techniques, allowing portable approximations across different GPU architectures. We conduct a comprehensive performance analysis of HPAC-Offload across GPU-accelerated HPC applications, revealing that AC techniques can significantly accelerate HPC applications (1.64x LULESH on AMD, 1.57x NVIDIA) with minimal quality loss (0.1%). Our analysis offers deep insights into the performance of GPU-based AC that guide the future development of AC algorithms and systems for these architectures.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper3
- ApproxTuner: a compiler and runtime system for adaptive approximationsHashim Sharif, Yifan Zhao, Maria Kotsifakou, Akash Kothari 等PPoPP 2021 · 被引用 22 次
- HPAC: evaluating approximate computing techniques on HPC OpenMP applicationsKonstantinos Parasyris, Giorgis Georgakoudis, Harshitha Menon, James Diffenderfer 等SC 2021 · 被引用 20 次
- Approximate Computing Through the Lens of Uncertainty QuantificationKonstantinos Parasyris, James Diffenderfer, Harshitha Menon, Ignacio Laguna 等SC 2022 · 被引用 5 次
相关 Paper
- Scalable Tuning of (OpenMP) GPU Applications via Kernel Record and ReplayKonstantinos Parasyris, Giorgis Georgakoudis, Esteban Rangel, Ignacio Laguna 等SC 2023 · 被引用 15 次
- CCAMP: an integrated translation and optimization framework for OpenACC and OpenMPJacob Lambert, Seyong Lee, Jeffrey S. Vetter, Allen D. MalonySC 2020 · 被引用 17 次
- Not All GPUs Are Created Equal: Characterizing Variability in Large-Scale, Accelerator-Rich SystemsPrasoon Sinha, Akhil Guliani, Rutwik Jain, Brandon Tran 等SC 2022 · 被引用 31 次
- Climbing the Summit and Pushing the Frontier of Mixed Precision Benchmarks at Extreme ScaleHao Lu, Michael A. Matheson, Vladyslav Oles, J. Austin Ellis 等SC 2022 · 被引用 8 次
- Profiling Hyperscale Big Data ProcessingAbraham Gonzalez, Aasheesh Kolli, Samira Manabi Khan, Sihang Liu 等ISCA 2023 · 被引用 30 次
