Historical Embedding-Guided Efficient Large-Scale Federated Graph Learning
Anran Li, Yuanyuan Chen, Jian Zhang, Mingfei Cheng, Yihao Huang, Yueming Wu, Anh Tuan Luu, Han Yu
摘要
Graph convolutional networks (GCNs) are promising for graph learning tasks. For privacy-preserving graph learning tasks involving distributed graph datasets, federated learning (FL)-based GCN (FedGCN) training is required. An important open challenge for FedGCN is scaling to large graphs, which typically incurs 1) high computation overhead for handling the explosively-increasing number of neighbors, and 2) high communication overhead of training GCNs involving multiple FL clients. Thus, neighbor sampling is being studied to enhance the scalability of FedGCNs. Existing FedGCN training techniques with neighbor sampling often produce extremely large communication and computation overhead and inaccurate node embeddings, leading to poor model performance. To bridge this gap, we propose the <u>Fed</u>erated <u>A</u>daptive <u>A</u>ttention-based <u>S</u>ampling (FedAAS) approach. It achieves substantial cost savings by efficiently leveraging historical embedding estimators and focusing the limited communication resources on transmitting the most influential neighbor node embeddings across FL clients. We further design an adaptive embedding synchronization scheme to optimize the efficiency and accuracy of FedAAS on large-scale datasets. Theoretical analysis shows that the approximation error induced by the staleness of historical embedding is upper bounded, and the model is guaranteed to converge in an efficient manner. Extensive experimental evaluation against four state-of-the-art baselines on six real-world graph datasets show that FedAAS achieves up to 5.12% higher test accuracy, while saving communication and computation costs by 95.11% and 94.76%, respectively.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper12
- Recipe for a General, Powerful, Scalable Graph TransformerLadislav Rampásek, Michael Galkin, Vijay Prakash Dwivedi, Anh Tuan Luu 等NeurIPS 2022 · 被引用 1,216 次
- GraphSAINT: Graph Sampling Based Inductive Learning MethodHanqing Zeng, Hongkuan Zhou, Ajitesh Srivastava, Rajgopal Kannan 等ICLR 2020 · 被引用 1,155 次
- Self-Supervised Graph Transformer on Large-Scale Molecular DataYu Rong, Yatao Bian, Tingyang Xu, Weiyang Xie 等NeurIPS 2020 · 被引用 1,113 次
- Federated Learning on Non-IID Data Silos: An Experimental StudyQinbin Li, Yiqun Diao, Quan Chen, Bingsheng HeICDE 2022 · 被引用 1,110 次
- Traffic Flow Prediction via Spatial Temporal Graph Neural NetworkXiaoyang Wang, Yao Ma, Yiqi Wang, Wei Jin 等WWW 2020 · 被引用 644 次
相关 Paper
- FedGCN: Convergence-Communication Tradeoffs in Federated Training of Graph Convolutional NetworksYuhang Yao, Weizhao Jin, Srivatsan Ravi, Carlee Joe-WongNeurIPS 2023 · 被引用 77 次
- GraphProxy: Communication-Efficient Federated Graph Learning with Adaptive ProxyJunyang Wang, Lan Zhang, Junhao Wang, Mu Yuan 等INFOCOM 2024 · 被引用 5 次
- ADGNN: Towards Scalable GNN Training with Aggregation-Difference Aware SamplingZhen Song, Yu Gu, Tianyi Li, Qing Sun 等SIGMOD 2024 · 被引用 9 次
- Dynamic Activation of Clients and Parameters for Federated Learning over Heterogeneous GraphsZishan Gu, Ke Zhang, Guangji Bai, Liang Chen 等ICDE 2023 · 被引用 10 次
- On Pipelined GCN with Communication-Efficient Sampling and Inclusion-Aware CachingShulin Wang, Qiang Yu, Xiong Wang, Yuqing Li 等INFOCOM 2024
