SCALE: A Structure-Centric Accelerator for Message Passing Graph Neural Networks
Lingxiang Yin, Sanjay Gandham, Mingjie Lin, Hao Zheng
摘要
Message passing paradigm has been widely used in developing complex Graph Neural Network (GNN) models, allowing for concise representations of edge and vertex-wise operations. Despite its pivotal role in theoretical advancement, the respective expression of edge and vertex operations, along with evolving GNN variants and datasets, has inevitably led to enormous computational complexity due to heterogeneous computation kernels. In particular, such inconsistent computation characteristics present new challenges in leveraging intermediate data reuse, ensuring both edge and vertex-wise workload balance, and sustaining system scalability. In this paper, we propose a structurecentric accelerator, SCALE, that can support a variety of message passing GNN models with improved parallelism, data reuse, and scalability. The central idea is to find latent similarities among GNN primitives such as shared dataflow structure, rather than strictly adhering to heterogeneous model structure. This serves as a hinge to homogenize inconsistencies in various GNN computation kernels. To accomplish this concept, SCALE consists of three unique designs, a novel systolic array-like architecture, a degree and vertex-aware scheduling, and a coherent dataflow tailored for fused graph and neural operations. The proposed systolic array-like architecture can support varying dataflows such as all-reduce, of distinct GNN operations improving parallelism, data reuse, and throughput. The degree and vertex-aware scheduling can remedy the workload imbalance encountered in vertex and edge-wise operations. Moreover, the proposed dataflow can unify the data movement of both graph and neural operators without extra communication and storage overheads. Our simulation results show that SCALE achieves 1.82× speedup and 38.9% energy reduction on average over the state-of-the-art GNN accelerators [1]–[4].
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper1
问问它们各自怎么用它相关 Paper
- EGMA: Enhancing Data Reuse and Workload Balancing in Message Passing GNN Acceleration via Gram Matrix OptimizationFangzhou Ye, Lingxiang Yin, Amir Ghazizadeh Ahsaei, Hao ZhengDAC 2024 · 被引用 5 次
- FlowGNN: A Dataflow Architecture for Real-Time Workload-Agnostic Graph Neural Network InferenceRishov Sarkar, Stefan Abi-Karam, Yuqi He, Lakshmi Sathidevi 等HPCA 2023 · 被引用 100 次
- I-DGNN: A Graph Dissimilarity-based Framework for Designing Scalable and Efficient DGNN AcceleratorsJiaqi Yang, Hao Zheng, Ahmed LouriHPCA 2025 · 被引用 3 次
- ReGNN: A Redundancy-Eliminated Graph Neural Networks AcceleratorCen Chen, Kenli Li, Yangfan Li, Xiaofeng ZouHPCA 2022 · 被引用 59 次
- uGrapher: High-Performance Graph Operator Computation via Unified Abstraction for Graph Neural NetworksYangjie Zhou, Jingwen Leng, Yaoxu Song, Shuwen Lu 等ASPLOS 2023 · 被引用 25 次
