Acceleration of fusion plasma turbulence simulations using the mixed-precision communication-avoiding krylov method
Yasuhiro Idomura, Takuya Ina, Yussuf Ali, Toshiyuki Imamura
摘要
The multi-scale full-f simulation of the next generation experimental fusion reactor ITER based on a five dimensional (5D) gyrokinetic model is one of the most computationally demanding problems in fusion science. In this work, a Gyrokinetic Toroidal 5D Eulerian code (GT5D) is accelerated by a new mixed-precision communication-avoiding (CA) Krylov method. The bottleneck of global collective communication on accelerated computing platforms is resolved using a CA Krylov method. In addition, a new FP16 preconditioner, which is designed using the new support for FP16 SIMD operations on A64FX, reduces both the number of iterations (halo data communication) and the computational cost. The performance of the proposed method for ITER size simulations with 0.1 trillion grids on 1,440 CPUs/GPUs on Fugaku and Summit shows 2.8× and 1.9× speedups respectively from the conventional non-CA Krylov method, and excellent strong scaling is obtained up to 5,760 CPUs/GPUs.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- 5 ExaFlop/s HPL-MxP Benchmark with Linear Scalability on the 40-Million-Core Sunway SupercomputerRongfen Lin, Xinhui Yuan, Wei Xue, Wanwang Yin 等SC 2023 · 被引用 11 次
- A Scalable Hybrid Total FETI Method for Massively Parallel FEM SimulationsKehao Lin, Chunbao Zhou, Yan Zeng, Ningming Nie 等PPoPP 2023 · 被引用 2 次
- Itoyori: Reconciling Global Address Space and Global Fork-Join Task ParallelismShumpei Shiina, Kenjiro TauraSC 2023 · 被引用 6 次
- Scaling Molecular Dynamics with ab initio Accuracy to 149 Nanoseconds per DayJianxiong Li, Boyang Li, Zhuoqiang Guo, Mingzhen Li 等SC 2024 · 被引用 9 次
- Petascale XCT: 3D image reconstruction with hierarchical communications on multi-GPU nodesMert Hidayetoglu, Tekin Bicer, Simon Garcia De Gonzalo, Bin Ren 等SC 2020 · 被引用 9 次
