SupeRBNN: Randomized Binary Neural Network Using Adiabatic Superconductor Josephson Devices
Zhengang Li, Geng Yuan, Tomoharu Yamauchi, Masoud Zabihi, Yanyue Xie, Peiyan Dong, Xulong Tang, Nobuyuki Yoshikawa, Devesh Tiwari, Yanzhi Wang, Olivia Chen
摘要
Adiabatic Quantum-Flux-Parametron (AQFP) is a superconducting logic with extremely high energy efficiency. By employing the distinct polarity of current to denote logic ‘0’ and ‘1’, AQFP devices serve as excellent carriers for binary neural network (BNN) computations. Although recent research has made initial strides toward developing an AQFP-based BNN accelerator, several critical challenges remain, preventing the design from being a comprehensive solution. In this paper, we propose SupeRBNN, an AQFP-based randomized BNN acceleration framework that leverages software-hardware co-optimization to eventually make the AQFP devices a feasible solution for BNN acceleration. Specifically, we investigate the randomized behavior of the AQFP devices and analyze the impact of crossbar size on current attenuation, subsequently formulating the current amplitude into the values suitable for use in BNN computation. To tackle the accumulation problem and improve overall hardware performance, we propose a stochastic computing-based accumulation module and a clocking scheme adjustment-based circuit optimization method. To effectively train the BNN models that are compatible with the distinctive characteristics of AQFP devices, we further propose a novel randomized BNN training solution that utilizes algorithm-hardware co-optimization, enabling simultaneous optimization of hardware configurations. In addition, we propose implementing batch normalization matching and the weight rectified clamp method to further improve the overall performance. We validate our SupeRBNN framework across various datasets and network architectures, comparing it with implementations based on different technologies, including CMOS, ReRAM, and superconducting RSFQ/ERSFQ. Experimental results demonstrate that our design achieves an energy efficiency of approximately 7.8 × 104 times higher than that of the ReRAM-based BNN framework while maintaining a similar level of model accuracy. Furthermore, when compared with superconductor-based counterparts, our framework demonstrates at least two orders of magnitude higher energy efficiency.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper6
- Rotated Binary Neural NetworkMingbao Lin, Rongrong Ji, Zihan Xu, Baochang Zhang 等NeurIPS 2020 · 被引用 161 次
- ReCU: Reviving the Dead Weights in Binary Neural NetworksZihan Xu, Mingbao Lin, Jianzhuang Liu, Jie Chen 等ICCV 2021 · 被引用 102 次
- Timely: Pushing Data Movements And Interfaces In Pim Accelerators Towards Local And In Time DomainWeitao Li, Pengfei Xu, Yang Zhao, Haitong Li 等ISCA 2020 · 被引用 86 次
- SuperNPU: An Extremely Fast Neural Processing Unit Using Superconducting Logic DevicesKoki Ishida, Ilkwon Byun, Ikki Nagaoka, Kosuke Fukumitsu 等MICRO 2020 · 被引用 66 次
- CryoCore: A Fast and Dense Processor Architecture for Cryogenic ComputingIlkwon Byun, Dongmoon Min, Gyu-hyeon Lee, Seongmin Na 等ISCA 2020 · 被引用 42 次
相关 Paper
- Unleashing the Potential of AQFP Logic Placement via Entanglement Entropy and ProjectionYinuo Bai, Enxin Yi, Wei W. Xing, Bei Yu 等DAC 2024 · 被引用 2 次
- TAAS: a timing-aware analytical strategy for AQFP-capable placement automationPeiyan Dong, Yanyue Xie, Hongjia Li, Mengshu Sun 等DAC 2022 · 被引用 7 次
- Beyond local optimality of buffer and splitter insertion for AQFP circuitsSiang-Yun Lee, Heinz Riener, Giovanni De MicheliDAC 2022 · 被引用 21 次
- RCGP: An Automatic Synthesis Framework for Reversible Quantum-Flux-Parametron Logic Circuits based on Efficient Cartesian Genetic ProgrammingRongliang Fu, Robert Wille, Tsung-Yi HoDAC 2024 · 被引用 4 次
- Sub-bit Neural Networks: Learning to Compress and Accelerate Binary Neural NetworksYikai Wang, Yi Yang, Fuchun Sun, Anbang YaoICCV 2021 · 被引用 18 次
