Lune

SC2020Top-tier venue

Reducing communication in graph neural network training

Alok Tripathy, Katherine A. Yelick, Aydin Buluç

2020Year
67Citations
25Top-tier citations

Abstract

Graph Neural Networks (GNNs) are powerful and flexible neural networks that use the naturally sparse connectivity information of the data. GNNs represent this connectivity as sparse matrices, which have lower arithmetic intensity and thus higher communication costs compared to dense matrices, making GNNs harder to scale to high concurrencies than convolutional or fully-connected neural networks.

We introduce a family of parallel algorithms for training GNNs and show that they can asymptotically reduce communication compared to previous parallel GNN training methods. We implement these algorithms, which are based on 1D, 1.5D, 2D, and 3D sparse-dense matrix multiplication, using torch.distributed on GPU-equipped clusters. Our algorithms optimize communication across the full GNN training pipeline. We train GNNs on over a hundred GPUs on multiple datasets, including a protein network with over a billion edges.

Index Terms-Graph neural networks, distributed training, communication-avoiding algorithms √ P ) fewer words than commonly utilized vertex-partitioning based approaches. The 3D algorithm we describe reduces the number

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext 48ee070e-610d-4e25-9722-86c29b6b9996

Cited by top-tier papers25

Ask how each one uses it

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines