Fograph: Enabling Real-Time Deep Graph Inference with Fog Computing
Liekang Zeng, Peng Huang, Ke Luo, Xiaoxi Zhang, Zhi Zhou, Xu Chen
摘要
Graph Neural Networks (GNNs) have gained growing interest in miscellaneous applications owing to their outstanding ability in extracting latent representation on graph structures. To render GNN-based service for IoT-driven smart applications, the traditional model serving paradigm resorts to the cloud by fully uploading the geo-distributed input data to the remote datacenter. However, our empirical measurements reveal the significant communication overhead of such cloud-based serving and highlight the profound potential in applying the emerging fog computing. To maximize the architectural benefits brought by fog computing, in this paper, we present Fograph, a novel distributed real-time GNN inference framework that leverages diverse resources of multiple fog nodes in proximity to IoT data sources. By introducing heterogeneity-aware execution planning and GNN-specific compression techniques, Fograph tailors its design to well accommodate the unique characteristics of GNN serving in fog environment. Prototype-based evaluation and case study demonstrate that Fograph significantly outperforms the state-of-the-art cloud serving and vanilla fog deployment by up to 5.39 × execution speedup and 6.84 × throughput improvement.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper2
- Asteroid: Resource-Efficient Hybrid Pipeline Parallelism for Collaborative DNN Training on Heterogeneous Edge DevicesShengyuan Ye, Liekang Zeng, Xiaowen Chu, Guoliang Xing 等MobiCom 2024 · 被引用 29 次
- λGrapher: A Resource-Efficient Serverless System for GNN Serving through Graph SharingHaichuan Hu, Fangming Liu, Qiangyu Pei, Yongjie Yuan 等WWW 2024 · 被引用 23 次
相关 Paper
- Graph Neural Networks Automated Design and Deployment on Device-Edge Co-Inference SystemsAo Zhou, Jianlei Yang, Tong Qiao, Yingjie Qi 等DAC 2024 · 被引用 5 次
- FLAG: An FPGA-Based System for Low-Latency GNN Inference Service Using Vector QuantizationYunki Han, Taehwan Kim, Jiwan Kim, Seohye Ha 等DAC 2025 · 被引用 1 次
- Distributed Inference Acceleration with Adaptive DNN Partitioning and OffloadingThaha Mohammed, Carlee Joe-Wong, Rohit Babbar, Mario Di FrancescoINFOCOM 2020 · 被引用 213 次
- EC-Graph: A Distributed Graph Neural Network System with Error-Compensated CompressionZhen Song, Yu Gu, Jianzhong Qi, Zhigang Wang 等ICDE 2022 · 被引用 14 次
- Hardware-Aware Graph Neural Network Automated Design for Edge Computing PlatformsAo Zhou, Jianlei Yang, Yingjie Qi, Yumeng Shi 等DAC 2023 · 被引用 15 次
