Input-Dependent Edge-Cloud Mapping of Recurrent Neural Networks Inference
Daniele Jahier Pagliari, Roberta Chiaro, Yukai Chen, Sara Vinco, Enrico Macii, Massimo Poncino
摘要
Given the computational complexity of Recurrent Neural Networks (RNNs) inference, IoT and mobile devices typically offload this task to the cloud. However, the execution time and energy consumption of RNN inference strongly depends on the length of the processed input. Therefore, considering also communication costs, it may be more convenient to process short input sequences locally and only offload long ones to the cloud. In this paper, we propose a low-overhead runtime tool that performs this choice automatically. Results based on real edge and cloud devices show that our method is able to simultaneously reduce the total execution time and energy consumption of the system compared to solutions that run RNN inference fully locally or fully in the cloud.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它相关 Paper
- SIEVE: Speculative Inference on the Edge with Versatile ExportationBabak Zamirai, Salar Latifi, Pedram Zamirai, Scott A. MahlkeDAC 2020 · 被引用 5 次
- Shiftry: RNN inference in 2KB of RAMAayan Kumar, Vivek Seshadri, Rahul SharmaOOPSLA 2020 · 被引用 18 次
- RTMobile: Beyond Real-Time Mobile Acceleration of RNNs for Speech RecognitionPeiyan Dong, Siyue Wang, Wei Niu, Chengming Zhang 等DAC 2020 · 被引用 50 次
- Extending the RISC-V ISA for Efficient RNN-based 5G Radio Resource ManagementRenzo Andri, Tomas Henriksson, Luca BeniniDAC 2020 · 被引用 10 次
- AutoScale: Energy Efficiency Optimization for Stochastic Edge Inference Using Reinforcement LearningYoung Geun Kim, Carole-Jean WuMICRO 2020 · 被引用 80 次
