SC2020Top-tier venue
Co-design for A64FX manycore processor and "Fugaku"
Mitsuhisa Sato, Yutaka Ishikawa, Hirofumi Tomita, Yuetsu Kodama, Tetsuya Odajima, Miwako Tsuji, Hisashi Yashiro, Masaki Aoki, Naoyuki Shida, Ikuo Miyoshi, Kouichi Hirai, Atsushi Furuya
Abstract
We have been carrying out the FLAGSHIP 2020 Project to develop the Japanese next-generation flagship supercomputer, the Post-K, recently named “Fugaku”. We have designed an original many core processor based on Armv8 instruction sets with the Scalable Vector Extension (SVE), an A64FX processor, as well as a system including interconnect and a storage subsystem with the industry partner, Fujitsu. The “co-design” of the system and applications is a key to making it power efficient and high performance. We determined many architectural parameters by reflecting an analysis of a set of target applications provided by applications teams. In this paper, we present the pragmatic practice of our co-design effort for “Fugaku”. As a result, the system has been proven to be a very power-efficient system, and it is confirmed that the performance of some target applications using the whole system is more than 100 times the performance of the K computer.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 15fc1af1-3f7d-4df8-bf1c-c104154bdfa2Cited by top-tier papers9
- LIBSHALOM: optimizing small and irregular-shaped matrix multiplications on ARMv8 multi-coresWeiling Yang, Jianbin Fang, Dezun Dong, Xing Su et al.SC 2021 · 40 citations
- High-Performance GPU-to-CPU Transpilation and Optimization via High-Level Parallel ConstructsWilliam S. Moses, Ivan R. Ivanov, Jens Domke, Toshio Endo et al.PPoPP 2023 · 27 citations
- autoGEMM: Pushing the Limits of Irregular Matrix Multiplication on Arm ArchitecturesDu Wu, Jintao Meng, Wenxi Zhu, Minwen Deng et al.SC 2024 · 14 citations
- MCBound: An Online Framework to Characterize and Classify Memory/Compute-bound HPC JobsFrancesco Antici, Andrea Bartolini, Zeynep Kiziltan, Özalp Babaoglu et al.SC 2024 · 10 citations
- Understanding Memory Failures on a Petascale Arm SystemKurt B. Ferreira, Scott Levy, Joshua Hemmert, Kevin T. PedrettiHPDC 2022 · 8 citations
Related papers
- Adaptable Register File Organization for Vector ProcessorsCristóbal Ramírez Lazo, Enrico Reggiani, Carlos Rojas Morales, Roger Figueras Bagué et al.HPCA 2022 · 7 citations
- Chronicles of astra: challenges and lessons from the first petascale arm supercomputerKevin T. Pedretti, Andrew J. Younge, Simon D. Hammond, James H. Laros III et al.SC 2020 · 13 citations
- Toward Sustainable HPC: In-Production Deployment of Incentive-Based Power Efficiency Mechanism on the Fugaku SupercomputerAna Luisa Veroneze Solórzano, Kento Sato, Keiji Yamamoto, Fumiyoshi Shoji et al.SC 2024 · 16 citations
- big.VLITTLE: On-Demand Data-Parallel Acceleration for Mobile Systems on ChipTuan Ta, Khalid Al-Hawaj, Nick Cebry, Yanghui Ou et al.MICRO 2022 · 10 citations
- APPEND: Rethinking ASIP Synthesis in the Era of AICangyuan Li, Ying Wang, Huawei Li, Yinhe HanDAC 2023 · 5 citations
