Lune

ICML2026顶会

Improving the Robustness-Utility Trade-off in Decentralized Learning over Sparse Networks

Yangnan Li, Xuanyu Cao, Shenghui Song

出版方
2026年份

摘要

Resilience against Byzantine attackers and faster convergence on sparse networks are critical for decentralized optimization, yet existing methods fail to achieve both simultaneously. Existing DSGD-based Byzantine-resilient methods suffer from high transient complexity of O((1−λ)−6)\mathcal{O}\left((1-\lambda)^{-6}\right), where 1−λ1-\lambda denotes the spectral gap of the network. While bias-correction methods such as Exact Diffusion can improve topology dependence, directly combining them with robust aggregators can lead to error accumulation. To address this issue, we introduce the scaled dual ascent (SDA) within the augmented Lagrangian framework for decentralized optimization, which mitigates error accumulation by scaling the dual update steps. Based on this, we propose BRED, which integrates Byzantine-robust Exact Diffusion with the SDA framework. We prove that BRED attains linear speedup, and achieves transient complexity of O((1−λ)−2)\mathcal{O}\left((1-\lambda)^{-2}\right) when the Byzantine fraction δ\delta is small. We further propose the momentum variant BRED-M, which reduces the Byzantine-affected transient complexity from O(δ2(1−λ)−6)\mathcal{O}\left(\delta^2(1-\lambda)^{-6}\right) to O(δ2(1−λ)−4)\mathcal{O}\left(\delta^2(1-\lambda)^{-4}\right). Empirical results on benchmark datasets demonstrate the efficacy of the proposed methods across diverse network topologies.

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

它引用的顶会 Paper4

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖