Albireo: Energy-Efficient Acceleration of Convolutional Neural Networks via Silicon Photonics
Kyle Shiflett, Avinash Karanth, Razvan C. Bunescu, Ahmed Louri
Abstract
With the end of Dennard scaling, highly-parallel and specialized hardware accelerators have been proposed to improve the throughput and energy-efficiency of deep neural network (DNN) models for various applications. However, collective data movement primitives such as multicast and broadcast that are required for multiply-and-accumulate (MAC) computation in DNN models are expensive, and require excessive energy and latency when implemented with electrical networks. This consequently limits the scalability and performance of electronic hardware accelerators. Emerging technology such as silicon photonics can inherently provide efficient implementation of multicast and broadcast operations, making photonics more amenable to exploit parallelism within DNN models. Moreover, when coupled with other unique features such as low energy consumption, high channel capacity with wavelength-division multiplexing (WDM), and high speed, silicon photonics could potentially provide a viable technology for scaling DNN acceleration.
In this paper, we propose Albireo, an analog photonic architecture for scaling DNN acceleration. By characterizing photonic devices such as microring resonators (MRRs) and Mach-Zehnder modulators (MZM) using photonic simulators, we develop realistic device models and outline their capability for system level acceleration. Using the device models, we develop an efficient broadcast combined with multicast data distribution by leveraging parameter sharing through unique WDM dot product processing. We evaluate the energy and throughput performance of Albireo on DNN models such as ResNet18, MobileNet and VGG16. When compared to current state-of-the-art electronic accelerators, Albireo increases throughput by 110 X, and improves energy-delay product (EDP) by an average of 74 X with current photonic devices. Furthermore, by considering moderate and aggressive photonic scaling, the proposed Albireo design shows that EDP can be reduced by at least 229 X.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext cdff7947-1e63-4710-8a59-d0596cf76bf3Cited by top-tier papers9
- Lightening-Transformer: A Dynamically-Operated Optically-Interconnected Photonic Transformer AcceleratorHanqing Zhu, Jiaqi Gu, Hanrui Wang, Zixuan Jiang et al.HPCA 2024 · 44 citations
- SPACX: Silicon Photonics-based Scalable Chiplet Accelerator for DNN InferenceYuan Li, Ahmed Louri, Avinash KaranthHPCA 2022 · 32 citations
- PhotoFourier: A Photonic Joint Transform Correlator-Based Neural Network AcceleratorShurui Li, Hangbo Yang, Chee Wei Wong, Volker J. Sorger et al.HPCA 2023 · 18 citations
- Mirage: An RNS-Based Photonic Accelerator for DNN TrainingCansu Demirkiran, Guowei Yang, Darius Bunandar, Ajay JoshiISCA 2024 · 16 citations
- Lightator: An Optical Near-Sensor Accelerator with Compressive Acquisition Enabling Versatile Image ProcessingMehrdad Morsali, Brendan Reidy, Deniz Najafi, Sepehr Tabrizchi et al.DAC 2024 · 15 citations
Builds on3
- SuperNPU: An Extremely Fast Neural Processing Unit Using Superconducting Logic DevicesKoki Ishida, Ilkwon Byun, Ikki Nagaoka, Kosuke Fukumitsu et al.MICRO 2020 · 66 citations
- PIXEL: Photonic Neural Network AcceleratorKyle Shiflett, Dylan Wright, Avinash Karanth, Ahmed LouriHPCA 2020 · 56 citations
- FReaC Cache: Folded-logic Reconfigurable Computing in the Last Level CacheAshutosh Dhar, Xiaohao Wang, Hubertus Franke, Jinjun Xiong et al.MICRO 2020 · 4 citations
Related papers
- Scaling Deep-Learning Inference with Chiplet-based Architecture and Photonic InterconnectsYuan Li, Ahmed Louri, Avinash KaranthDAC 2021 · 21 citations
- Towards Memory-Efficient Neural Networks via Multi-Level in situ GenerationJiaqi Gu, Hanqing Zhu, Chenghao Feng, Mingjie Liu et al.ICCV 2021 · 4 citations
- CrossLight: A Cross-Layer Optimized Silicon Photonic Neural Network AcceleratorFebin Sunny, Asif Mirza, Mahdi Nikdast, Sudeep PasrichaDAC 2021 · 92 citations
- SiP-ML: high-bandwidth optical network interconnects for machine learning trainingMehrdad Khani Shirkoohi, Manya Ghobadi, Mohammad Alizadeh, Ziyi Zhu et al.SIGCOMM 2021 · 94 citations
- Balancing Efficiency and Flexibility for DNN Acceleration via Temporal GPU-Systolic Array IntegrationCong Guo, Yangjie Zhou, Jingwen Leng, Yuhao Zhu et al.DAC 2020 · 35 citations
