Cross-Gradient Aggregation for Decentralized Learning from Non-IID Data
Yasaman Esfandiari, Sin Yong Tan, Zhanhong Jiang, Aditya Balu, Ethan Herron, Chinmay Hegde, Soumik Sarkar
摘要
Decentralized learning enables a group of collaborative agents to learn models using a distributed dataset without the need for a central parameter server. Recently, decentralized learning algorithms have demonstrated state-of-the-art results on benchmark data sets, comparable with centralized algorithms. However, the key assumption to achieve competitive performance is that the data is independently and identically distributed (IID) among the agents which, in real-life applications, is often not applicable. Inspired by ideas from continual learning, we propose Cross-Gradient Aggregation (CGA), a novel decentralized learning algorithm where (i) each agent aggregates cross-gradient information, i.e., derivatives of its model with respect to its neighbors' datasets, and (ii) updates its model using a projected gradient based on quadratic programming (QP). We theoretically analyze the convergence characteristics of CGA and demonstrate its efficiency on non-IID data distributions sampled from the MNIST and CIFAR-10 datasets. Our empirical comparisons show superior learning performance of CGA over existing state-of-the-art decentralized learning algorithms, as well as maintaining the improved performance under information compression to reduce peer-to-peer communication overhead. The code is available here on GitHub.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Learning to Collaborate in Decentralized Learning of Personalized ModelsShuangtong Li, Tianyi Zhou, Xinmei Tian, Dacheng TaoCVPR 2022 · 被引用 41 次
- Global Update Tracking: A Decentralized Learning Algorithm for Heterogeneous DataSai Aparna Aketi, Abolfazl Hashemi, Kaushik RoyNeurIPS 2023 · 被引用 23 次
- Decentralized Dynamic Cooperation of Personalized Models for Federated Continual LearningDanni Yang, Zhikang Chen, Sen Cui, Mengyue Yang 等NeurIPS 2025 · 被引用 2 次
- DIMAT: Decentralized Iterative Merging-And-Training for Deep Learning ModelsNastaran Saadati, Minh Pham, Nasla Saleem, Joshua R. Waite 等CVPR 2024 · 被引用 2 次
- GradMA: A Gradient-Memory-based Accelerated Federated Learning with Alleviated Catastrophic ForgettingKangyang Luo, Xiang Li, Yunshi Lan, Ming GaoCVPR 2023
它引用的顶会 Paper6
- On the Convergence of FedAvg on Non-IID DataXiang Li, Kaixuan Huang, Wenhao Yang, Shusen Wang 等ICLR 2020 · 被引用 2,930 次
- The Non-IID Data Quagmire of Decentralized Machine LearningKevin Hsieh, Amar Phanishayee, Onur Mutlu, Phillip B. GibbonsICML 2020 · 被引用 672 次
- A Unified Theory of Decentralized SGD with Changing Topology and Local UpdatesAnastasia Koloskova, Nicolas Loizou, Sadra Boreiri, Martin Jaggi 等ICML 2020 · 被引用 623 次
- Decentralized Deep Learning with Arbitrary Communication CompressionAnastasia Koloskova, Tao Lin, Sebastian U. Stich, Martin JaggiICLR 2020 · 被引用 263 次
- Moniqua: Modulo Quantized Communication in Decentralized SGDYucheng Lu, Christopher De SaICML 2020 · 被引用 53 次
相关 Paper
- Robust Distributed Gradient Aggregation Using Projections onto Gradient ManifoldsKwang In KimAAAI 2024
- Decentralized Learning with Multi-Headed DistillationAndrey Zhmoginov, Mark Sandler, Nolan Miller, Gus Kristiansen 等CVPR 2023
- Robust Combination of Distributed Gradients Under Adversarial PerturbationsKwang In KimCVPR 2022 · 被引用 2 次
- Decentralized Sporadic Federated Learning: A Unified Algorithmic Framework with Convergence GuaranteesShahryar Zehtabi, Dong-Jun Han, Rohit Parasnis, Seyyedali Hosseinalipour 等ICLR 2025
- Optimal Complexity in Decentralized TrainingYucheng Lu, Christopher De SaICML 2021 · 被引用 95 次
