Graph Neural Networks Automated Design and Deployment on Device-Edge Co-Inference Systems
Ao Zhou, Jianlei Yang, Tong Qiao, Yingjie Qi, Zhi Yang, Weisheng Zhao, Chunming Hu
Abstract
The key to device-edge co-inference paradigm is to partition models into computation-friendly and computation-intensive parts across the device and the edge, respectively. However, for Graph Neural Networks (GNNs), we find that simply partitioning without altering their structures can hardly achieve the full potential of the co-inference paradigm due to various computational-communication overheads of GNN operations over heterogeneous devices. We present GCoDE, the first automatic framework for GNN that innovatively Co-designs the architecture search and the mapping of each operation on Device-Edge hierarchies. GCoDE abstracts the device communication process into an explicit operation and fuses the search of architecture and the operations mapping in a unified space for joint-optimization. Also, the performance-awareness approach, utilized in the constraint-based search process of GCoDE, enables effective evaluation of architecture efficiency in diverse heterogeneous systems. We implement the co-inference engine and runtime dispatcher in GCoDE to enhance the deployment efficiency. Experimental results show that GCoDE can achieve up to 44.9× speedup and 98.2% energy reduction compared to existing approaches across various applications and system configurations.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1fedddd4-6435-4139-8f8a-6632e2ecadb1Builds on5
- Evaluating The Search Phase of Neural Architecture SearchKaicheng Yu, Christian Sciuto, Martin Jaggi, Claudiu Musat et al.ICLR 2020 · 370 citations
- Towards Efficient Graph Convolutional Networks for Point Cloud HandlingYawei Li, He Chen, Zhaopeng Cui, Radu Timofte et al.ICCV 2021 · 32 citations
- Hardware-Aware Graph Neural Network Automated Design for Edge Computing PlatformsAo Zhou, Jianlei Yang, Yingjie Qi, Yumeng Shi et al.DAC 2023 · 15 citations
- LENS: Layer Distribution Enabled Neural Architecture Search in Edge-Cloud HierarchiesMohanad Odema, Nafiul Rashid, Berken Utku Demirel, Mohammad Abdullah Al FaruqueDAC 2021 · 8 citations
- Multivariate, Multi-Frequency and Multimodal: Rethinking Graph Neural Networks for Emotion Recognition in ConversationFeiyu Chen, Jie Shao, Shuyuan Zhu, Heng Tao ShenCVPR 2023
Related papers
- Fograph: Enabling Real-Time Deep Graph Inference with Fog ComputingLiekang Zeng, Peng Huang, Ke Luo, Xiaoxi Zhang et al.WWW 2022 · 22 citations
- EdgeNN: Efficient Neural Network Inference for CPU-GPU Integrated Edge DevicesChenyang Zhang, Feng Zhang, Kuangyu Chen, Mingjun Chen et al.ICDE 2023 · 16 citations
- MEGA: A Memory-Efficient GNN Accelerator Exploiting Degree-Aware Mixed-Precision QuantizationZeyu Zhu, Fanrong Li, Gang Li, Zejian Liu et al.HPCA 2024 · 28 citations
- GSPO: A Graph Substitution and Parallelization Joint Optimization Framework for DNN InferenceZheng Xu, Xu Dai, Shaojun Wei, Shouyi Yin et al.DAC 2024 · 3 citations
- Hardware Acceleration of Graph Neural NetworksAdam Auten, Matthew Tomei, Rakesh KumarDAC 2020 · 108 citations
