Lune

DAC2022顶会

GNNIE: GNN inference engine with load-balancing and graph-specific caching

Sudipta Mondal, Susmita Dey Manasi, Kishor Kunal, Ramprasath S, Sachin S. Sapatnekar

2022年份
20被引次数
5顶会引用

摘要

Graph neural networks (GNN) analysis engines are vital for real-world problems that use large graph models. Challenges for a GNN hardware platform include the ability to (a) host a variety of GNNs, (b) handle high sparsity in input vertex feature vectors and the graph adjacency matrix and the accompanying random memory access patterns, and (c) maintain load-balanced computation in the face of uneven workloads, induced by high sparsity and power-law vertex degree distributions. This paper proposes GNNIE, an accelerator designed to run a broad range of GNNs. It tackles workload imbalance by (i) splitting vertex feature operands into blocks, (ii) reordering and redistributing computations, (iii) using a novel flexible MAC architecture. It adopts a graph-specific, degree-aware caching policy that is well suited to real-world graph characteristics. The policy enhances on-chip data reuse and avoids random memory access to DRAM.

GNNIE achieves average speedups of 21233× over a CPU and 699× over a GPU over multiple datasets on graph attention networks (GATs), graph convolutional networks (GCNs), Graph-SAGE, GINConv, and DiffPool. Compared to prior approaches, GNNIE achieves an average speedup of 35× over HyGCN (which cannot implement GATs) for GCN, GraphSAGE, and GINConv, and, using 3.4× fewer processing units, an average speedup of 2.1× over AWB-GCN (which runs only GCNs).

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

引用它的顶会 Paper5

问问它们各自怎么用它

它引用的顶会 Paper7

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖