Lune

HPDC2024Top-tier venue

ADTopk: All-Dimension Top-k Compression for High-Performance Data-Parallel DNN Training

Zhangqiang Ming, Yuchong Hu, Wenxiang Zhou, Xinjue Zheng, Chenxuan Yao, Dan Feng

2024Year
5Citations

Abstract

Data-parallel deep neural networks (DNN) training systems deployed across nodes have been widely used in various domains, while the system performance is often bottlenecked by the communication overhead among workers for synchronizing gradients. Top-k sparsification compression is the de facto approach to alleviate the communication bottleneck, which truncates the gradient to its largest k elements before sending it to other nodes.

Ask about this paper

Ask your agent about it.

Lune has read the top-tier papers around this one, so every answer names the papers it rests on.

Questions to start from

Your agent calls

Lunesearch_papers

Ask in Lune

Free to start. No credit card required.

lune papers get 8872f267-2041-41ea-aa42-5cc433fd50a7

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines