AWB-GCN: A Graph Convolutional Network Accelerator with Runtime Workload Rebalancing
Tong Geng, Ang Li, Runbin Shi, Chunshu Wu, Tianqi Wang, Yanfei Li, Pouya Haghi, Antonino Tumeo, Shuai Che, Steven K. Reinhardt, Martin C. Herbordt
摘要
Deep learning systems have been successfully applied to Euclidean data such as images, video, and audio. In many applications, however, information and their relationships are better expressed with graphs. Graph Convolutional Networks (GCNs) appear to be a promising approach to efficiently learn from graph data structures, having shown advantages in many critical applications. As with other deep learning modalities, hardware acceleration is critical. The challenge is that real-world graphs are often extremely large and unbalanced; this poses significant performance demands and design challenges.
In this paper, we propose Autotuning-Workload-Balancing GCN (AWB-GCN) to accelerate GCN inference. To address the issue of workload imbalance in processing real-world graphs, three hardware-based autotuning techniques are proposed: dynamic distribution smoothing, remote switching, and row remapping. In particular, AWB-GCN continuously monitors the sparse graph pattern, dynamically adjusts the workload distribution among a large number of processing elements (up to 4K PEs), and, after converging, reuses the ideal configuration. Evaluation is performed using an Intel D5005 FPGA with five commonly-used datasets. Results show that 4K-PE AWB-GCN can significantly elevate PE utilization by 7.7× on average and demonstrate considerable performance speedups over CPUs (3255×), GPUs (80.3×), and a prior GCN accelerator (5.1×).
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper33
- A Unified Lottery Ticket Hypothesis for Graph Neural NetworksTianlong Chen, Yongduo Sui, Xuxi Chen, Aston Zhang 等ICML 2021 · 被引用 208 次
- GCNAX: A Flexible and Energy-efficient Accelerator for Graph Convolutional Neural NetworksJiajun Li, Ahmed Louri, Avinash Karanth, Razvan C. BunescuHPCA 2021 · 被引用 147 次
- I-GCN: A Graph Convolutional Network Accelerator with Runtime Locality Enhancement through IslandizationTong Geng, Chunshu Wu, Yongan Zhang, Cheng Tan 等MICRO 2021 · 被引用 138 次
- FlowGNN: A Dataflow Architecture for Real-Time Workload-Agnostic Graph Neural Network InferenceRishov Sarkar, Stefan Abi-Karam, Yuqi He, Lakshmi Sathidevi 等HPCA 2023 · 被引用 100 次
- GCoD: Graph Convolutional Network Acceleration via Dedicated Algorithm and Accelerator Co-DesignHaoran You, Tong Geng, Yongan Zhang, Ang Li 等HPCA 2022 · 被引用 66 次
它引用的顶会 Paper3
- SIGMA: A Sparse and Irregular GEMM Accelerator with Flexible Interconnects for DNN TrainingEric Qin, Ananda Samajdar, Hyoukjun Kwon, Vineet Nadella 等HPCA 2020 · 被引用 490 次
- HyGCN: A GCN Accelerator with Hybrid ArchitectureMingyu Yan, Lei Deng, Xing Hu, Ling Liang 等HPCA 2020 · 被引用 338 次
- ALRESCHA: A Lightweight Reconfigurable Sparse-Computation AcceleratorBahar Asgari, Ramyad Hadidi, Tushar Krishna, Hyesoon Kim 等HPCA 2020 · 被引用 63 次
相关 Paper
- GNNIE: GNN inference engine with load-balancing and graph-specific cachingSudipta Mondal, Susmita Dey Manasi, Kishor Kunal, Ramprasath S 等DAC 2022 · 被引用 20 次
- Accelerating Graph Convolutional Networks Using Crossbar-based Processing-In-Memory ArchitecturesYu Huang, Long Zheng, Pengcheng Yao, Qinggang Wang 等HPCA 2022 · 被引用 63 次
- AutoGNN: End-to-End Hardware-Driven Graph Preprocessing for Enhanced GNN PerformanceSeungkwan Kang, Seungjun Lee, Donghyun Gouk, Miryeong Kwon 等HPCA 2026
- An Efficient Hardware Accelerator Design for Dynamic Graph Convolutional Network (DGCN) InferenceYingnan Zhao, Ke Wang, Jiaqi Yang, Ahmed LouriDAC 2024 · 被引用 3 次
- SparseWeaver: Converting Sparse Operations as Dense Operations on GPUs for Graph WorkloadsShinnung Jeong, Liam Paul Cooper, Ju Min Lee, Heelim Choi 等HPCA 2025 · 被引用 2 次
