Co-design for A64FX manycore processor and "Fugaku"
Mitsuhisa Sato, Yutaka Ishikawa, Hirofumi Tomita, Yuetsu Kodama, Tetsuya Odajima, Miwako Tsuji, Hisashi Yashiro, Masaki Aoki, Naoyuki Shida, Ikuo Miyoshi, Kouichi Hirai, Atsushi Furuya
摘要
We have been carrying out the FLAGSHIP 2020 Project to develop the Japanese next-generation flagship supercomputer, the Post-K, recently named “Fugaku”. We have designed an original many core processor based on Armv8 instruction sets with the Scalable Vector Extension (SVE), an A64FX processor, as well as a system including interconnect and a storage subsystem with the industry partner, Fujitsu. The “co-design” of the system and applications is a key to making it power efficient and high performance. We determined many architectural parameters by reflecting an analysis of a set of target applications provided by applications teams. In this paper, we present the pragmatic practice of our co-design effort for “Fugaku”. As a result, the system has been proven to be a very power-efficient system, and it is confirmed that the performance of some target applications using the whole system is more than 100 times the performance of the K computer.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper9
- LIBSHALOM: optimizing small and irregular-shaped matrix multiplications on ARMv8 multi-coresWeiling Yang, Jianbin Fang, Dezun Dong, Xing Su 等SC 2021 · 被引用 40 次
- High-Performance GPU-to-CPU Transpilation and Optimization via High-Level Parallel ConstructsWilliam S. Moses, Ivan R. Ivanov, Jens Domke, Toshio Endo 等PPoPP 2023 · 被引用 27 次
- autoGEMM: Pushing the Limits of Irregular Matrix Multiplication on Arm ArchitecturesDu Wu, Jintao Meng, Wenxi Zhu, Minwen Deng 等SC 2024 · 被引用 14 次
- MCBound: An Online Framework to Characterize and Classify Memory/Compute-bound HPC JobsFrancesco Antici, Andrea Bartolini, Zeynep Kiziltan, Özalp Babaoglu 等SC 2024 · 被引用 10 次
- Understanding Memory Failures on a Petascale Arm SystemKurt B. Ferreira, Scott Levy, Joshua Hemmert, Kevin T. PedrettiHPDC 2022 · 被引用 8 次
相关 Paper
- Adaptable Register File Organization for Vector ProcessorsCristóbal Ramírez Lazo, Enrico Reggiani, Carlos Rojas Morales, Roger Figueras Bagué 等HPCA 2022 · 被引用 7 次
- Chronicles of astra: challenges and lessons from the first petascale arm supercomputerKevin T. Pedretti, Andrew J. Younge, Simon D. Hammond, James H. Laros III 等SC 2020 · 被引用 13 次
- Toward Sustainable HPC: In-Production Deployment of Incentive-Based Power Efficiency Mechanism on the Fugaku SupercomputerAna Luisa Veroneze Solórzano, Kento Sato, Keiji Yamamoto, Fumiyoshi Shoji 等SC 2024 · 被引用 16 次
- big.VLITTLE: On-Demand Data-Parallel Acceleration for Mobile Systems on ChipTuan Ta, Khalid Al-Hawaj, Nick Cebry, Yanghui Ou 等MICRO 2022 · 被引用 10 次
- APPEND: Rethinking ASIP Synthesis in the Era of AICangyuan Li, Ying Wang, Huawei Li, Yinhe HanDAC 2023 · 被引用 5 次
