SuperNPU: An Extremely Fast Neural Processing Unit Using Superconducting Logic Devices
Koki Ishida, Ilkwon Byun, Ikki Nagaoka, Kosuke Fukumitsu, Masamitsu Tanaka, Satoshi Kawakami, Teruo Tanimoto, Takatsugu Ono, Jangwoo Kim, Koji Inoue
摘要
Superconductor single-flux-quantum (SFQ) logic family has been recognized as a highly promising solution for the post-Moore's era, thanks to its ultra-fast and low-power switching characteristics. Therefore, researchers have made a tremendous amount of effort in various aspects to promote the technology and automate its circuit design process (e.g., low-cost fabrication, design tool development). However, there has been no progress in designing a convincing SFQ-based architectural unit due to the architects' lack of understanding of the technology's potentials and limitations at the architecture level. In this paper, we present how to architect an SFQ-based architectural unit by providing design principles with an extreme-performance neural processing unit (NPU). To achieve the goal, we first implement an architecture-level simulator to model an SFQ-based NPU accurately. We validate this model using our die-level prototypes, design tools, and logic cell library. This simulator accurately measures the NPU's performance, power consumption, area, and cooling overheads. Next, driven by the modeling, we identify key architectural challenges for designing a performance-effective SFQ-based NPU (e.g., expensive on-chip data movements and buffering). Lastly, we present SuperNPU, our example SFQ-based NPU architecture, which effectively resolves the challenges. Our evaluation shows that the proposed design outperforms a conventional state-of-the-art NPU by 23 times. With free cooling provided as done in quantum computing, the performance per chip power increases up to 490 times. Our methodology can also be applied to other architecture designs with SFQ-friendly characteristics.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- SpAtten: Efficient Sparse Attention Architecture with Cascade Token and Head PruningHanrui Wang, Zhekai Zhang, Song HanHPCA 2021 · 被引用 412 次
- Albireo: Energy-Efficient Acceleration of Convolutional Neural Networks via Silicon PhotonicsKyle Shiflett, Avinash Karanth, Razvan C. Bunescu, Ahmed LouriISCA 2021 · 被引用 44 次
- DigiQ: A Scalable Digital Controller for Quantum Computers Using SFQ LogicMohammad Reza Jokar, Richard Rines, Ghasem Pasandi, Haolin Cong 等HPCA 2022 · 被引用 37 次
- SMART: A Heterogeneous Scratchpad Memory Architecture for Superconductor SFQ-based Systolic CNN AcceleratorsFarzaneh Zokaee, Lei JiangMICRO 2021 · 被引用 20 次
- SupeRBNN: Randomized Binary Neural Network Using Adiabatic Superconductor Josephson DevicesZhengang Li, Geng Yuan, Tomoharu Yamauchi, Masoud Zabihi 等MICRO 2023 · 被引用 7 次
它引用的顶会 Paper2
相关 Paper
- SuperSFQ: A Hardware Design to Realize High-Frequency Superconducting ProcessorsJunhyuk Choi, Juwon Hong, Junpyo Kim, Jungmin Cho 等MICRO 2025 · 被引用 2 次
- SuperCore: An Ultra-Fast Superconducting Processor for Cryogenic ApplicationsJunhyuk Choi, Ilkwon Byun, Juwon Hong, Dongmoon Min 等MICRO 2024 · 被引用 9 次
- SUSHI: Ultra-High-Speed and Ultra-Low-Power Neuromorphic Chip Using Superconducting Single-Flux-Quantum CircuitsZeshi Liu, Shuo Chen, Peiyao Qu, Huanli Liu 等MICRO 2023 · 被引用 11 次
- Temporal and SFQ pulse-streams encoding for area-efficient superconducting acceleratorsPatricia Gonzalez-Guerrero, Meriam Gay Bautista, Darren Lyles, George MichelogiannakisASPLOS 2022 · 被引用 18 次
- SuperBP: Design Space Exploration of Perceptron-Based Branch Predictors for Superconducting CPUsHaipeng Zha, Swamit Tannu, Murali AnnavaramMICRO 2023 · 被引用 1 次
