GCoD: Graph Convolutional Network Acceleration via Dedicated Algorithm and Accelerator Co-Design
Haoran You, Tong Geng, Yongan Zhang, Ang Li, Yingyan Lin
摘要
Graph Convolutional Networks (GCNs) have emerged as the state-of-the-art graph learning model. However, it can be notoriously challenging to inference GCNs over large graph datasets, limiting their application to large real-world graphs and hindering the exploration of deeper and more sophisticated GCN graphs. This is because real-world graphs can be extremely large and sparse. Furthermore, the node degree of GCNs tends to follow the power-law distribution and therefore have highly irregular adjacency matrices, resulting in prohibitive inefficiencies in both data processing and movement and thus substantially limiting the achievable GCN acceleration efficiency. To this end, this paper proposes a GCN algorithm and accelerator Co-Design framework dubbed GCoD which can largely alleviate the aforementioned GCN irregularity and boost GCNs' inference efficiency. Specifically, on the algorithm level, GCoD integrates a split and conquer GCN training strategy that polarizes the graphs to be either denser or sparser in local neighborhoods without compromising the model accuracy, resulting in graph adjacency matrices that (mostly) have merely two levels of workload and enjoys largely enhanced regularity and thus ease of acceleration. On the hardware level, we further develop a dedicated twopronged accelerator with a separated engine to process each of the aforementioned denser and sparser workloads, further boosting the overall utilization and acceleration efficiency. Extensive experiments and ablation studies validate that our GCoD consistently reduces the number of off-chip accesses, leading to speedups 15286×, 294×, 7.8×, and 2.5× as compared to CPUs, GPUs, and prior-art GCN accelerators including HyGCN and AWB-GCN, respectively, while maintaining or even improving the task accuracy. Additionally, we visualize GCoD trained graph adjacency matrices for a better understanding of its advantages. Codes are available at https: //github.com/RICE-EIC/GCoD.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- FlowGNN: A Dataflow Architecture for Real-Time Workload-Agnostic Graph Neural Network InferenceRishov Sarkar, Stefan Abi-Karam, Yuqi He, Lakshmi Sathidevi 等HPCA 2023 · 被引用 100 次
- Do We Need Anisotropic Graph Neural Networks?Shyam A. Tailor, Felix L. Opolka, Pietro Liò, Nicholas Donald LaneICLR 2022 · 被引用 46 次
- Ising-Traffic: Using Ising Machine Learning to Predict Traffic Congestion under UncertaintyZhenyu Pan, Anshujit Sharma, Jerry Yao-Chieh Hu, Zhuo Liu 等AAAI 2023 · 被引用 43 次
- MaxK-GNN: Extremely Fast GPU Kernel Design for Accelerating Graph Neural Networks TrainingHongwu Peng, Xi Xie, Kaustubh Shivdikar, Md Amit Hasan 等ASPLOS 2024 · 被引用 32 次
- MEGA: A Memory-Efficient GNN Accelerator Exploiting Degree-Aware Mixed-Precision QuantizationZeyu Zhu, Fanrong Li, Gang Li, Zejian Liu 等HPCA 2024 · 被引用 28 次
它引用的顶会 Paper10
- Open Graph Benchmark: Datasets for Machine Learning on GraphsWeihua Hu, Matthias Fey, Marinka Zitnik, Yuxiao Dong 等NeurIPS 2020 · 被引用 3,935 次
- Learning Graph Convolutional Network for Skeleton-Based Human Action Recognition by Neural SearchingWei Peng, Xiaopeng Hong, Haoyu Chen, Guoying ZhaoAAAI 2020 · 被引用 362 次
- HyGCN: A GCN Accelerator with Hybrid ArchitectureMingyu Yan, Lei Deng, Xing Hu, Ling Liang 等HPCA 2020 · 被引用 338 次
- Robust Graph Representation Learning via Neural SparsificationCheng Zheng, Bo Zong, Wei Cheng, Dongjin Song 等ICML 2020 · 被引用 330 次
- AWB-GCN: A Graph Convolutional Network Accelerator with Runtime Workload RebalancingTong Geng, Ang Li, Runbin Shi, Chunshu Wu 等MICRO 2020 · 被引用 299 次
相关 Paper
- I-GCN: A Graph Convolutional Network Accelerator with Runtime Locality Enhancement through IslandizationTong Geng, Chunshu Wu, Yongan Zhang, Cheng Tan 等MICRO 2021 · 被引用 138 次
- GCNAX: A Flexible and Energy-efficient Accelerator for Graph Convolutional Neural NetworksJiajun Li, Ahmed Louri, Avinash Karanth, Razvan C. BunescuHPCA 2021 · 被引用 147 次
- Accelerating Graph Convolutional Networks Using Crossbar-based Processing-In-Memory ArchitecturesYu Huang, Long Zheng, Pengcheng Yao, Qinggang Wang 等HPCA 2022 · 被引用 63 次
- GNNIE: GNN inference engine with load-balancing and graph-specific cachingSudipta Mondal, Susmita Dey Manasi, Kishor Kunal, Ramprasath S 等DAC 2022 · 被引用 20 次
- GNNerator: A Hardware/Software Framework for Accelerating Graph Neural NetworksJacob R. Stevens, Dipankar Das, Sasikanth Avancha, Bharat Kaul 等DAC 2021 · 被引用 22 次
