Albireo: Energy-Efficient Acceleration of Convolutional Neural Networks via Silicon Photonics
Kyle Shiflett, Avinash Karanth, Razvan C. Bunescu, Ahmed Louri
摘要
With the end of Dennard scaling, highly-parallel and specialized hardware accelerators have been proposed to improve the throughput and energy-efficiency of deep neural network (DNN) models for various applications. However, collective data movement primitives such as multicast and broadcast that are required for multiply-and-accumulate (MAC) computation in DNN models are expensive, and require excessive energy and latency when implemented with electrical networks. This consequently limits the scalability and performance of electronic hardware accelerators. Emerging technology such as silicon photonics can inherently provide efficient implementation of multicast and broadcast operations, making photonics more amenable to exploit parallelism within DNN models. Moreover, when coupled with other unique features such as low energy consumption, high channel capacity with wavelength-division multiplexing (WDM), and high speed, silicon photonics could potentially provide a viable technology for scaling DNN acceleration.
In this paper, we propose Albireo, an analog photonic architecture for scaling DNN acceleration. By characterizing photonic devices such as microring resonators (MRRs) and Mach-Zehnder modulators (MZM) using photonic simulators, we develop realistic device models and outline their capability for system level acceleration. Using the device models, we develop an efficient broadcast combined with multicast data distribution by leveraging parameter sharing through unique WDM dot product processing. We evaluate the energy and throughput performance of Albireo on DNN models such as ResNet18, MobileNet and VGG16. When compared to current state-of-the-art electronic accelerators, Albireo increases throughput by 110 X, and improves energy-delay product (EDP) by an average of 74 X with current photonic devices. Furthermore, by considering moderate and aggressive photonic scaling, the proposed Albireo design shows that EDP can be reduced by at least 229 X.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- Lightening-Transformer: A Dynamically-Operated Optically-Interconnected Photonic Transformer AcceleratorHanqing Zhu, Jiaqi Gu, Hanrui Wang, Zixuan Jiang 等HPCA 2024 · 被引用 44 次
- SPACX: Silicon Photonics-based Scalable Chiplet Accelerator for DNN InferenceYuan Li, Ahmed Louri, Avinash KaranthHPCA 2022 · 被引用 32 次
- PhotoFourier: A Photonic Joint Transform Correlator-Based Neural Network AcceleratorShurui Li, Hangbo Yang, Chee Wei Wong, Volker J. Sorger 等HPCA 2023 · 被引用 18 次
- Mirage: An RNS-Based Photonic Accelerator for DNN TrainingCansu Demirkiran, Guowei Yang, Darius Bunandar, Ajay JoshiISCA 2024 · 被引用 16 次
- Lightator: An Optical Near-Sensor Accelerator with Compressive Acquisition Enabling Versatile Image ProcessingMehrdad Morsali, Brendan Reidy, Deniz Najafi, Sepehr Tabrizchi 等DAC 2024 · 被引用 15 次
它引用的顶会 Paper3
- SuperNPU: An Extremely Fast Neural Processing Unit Using Superconducting Logic DevicesKoki Ishida, Ilkwon Byun, Ikki Nagaoka, Kosuke Fukumitsu 等MICRO 2020 · 被引用 66 次
- PIXEL: Photonic Neural Network AcceleratorKyle Shiflett, Dylan Wright, Avinash Karanth, Ahmed LouriHPCA 2020 · 被引用 56 次
- FReaC Cache: Folded-logic Reconfigurable Computing in the Last Level CacheAshutosh Dhar, Xiaohao Wang, Hubertus Franke, Jinjun Xiong 等MICRO 2020 · 被引用 4 次
相关 Paper
- Scaling Deep-Learning Inference with Chiplet-based Architecture and Photonic InterconnectsYuan Li, Ahmed Louri, Avinash KaranthDAC 2021 · 被引用 21 次
- Towards Memory-Efficient Neural Networks via Multi-Level in situ GenerationJiaqi Gu, Hanqing Zhu, Chenghao Feng, Mingjie Liu 等ICCV 2021 · 被引用 4 次
- CrossLight: A Cross-Layer Optimized Silicon Photonic Neural Network AcceleratorFebin Sunny, Asif Mirza, Mahdi Nikdast, Sudeep PasrichaDAC 2021 · 被引用 92 次
- SiP-ML: high-bandwidth optical network interconnects for machine learning trainingMehrdad Khani Shirkoohi, Manya Ghobadi, Mohammad Alizadeh, Ziyi Zhu 等SIGCOMM 2021 · 被引用 94 次
- Balancing Efficiency and Flexibility for DNN Acceleration via Temporal GPU-Systolic Array IntegrationCong Guo, Yangjie Zhou, Jingwen Leng, Yuhao Zhu 等DAC 2020 · 被引用 35 次
