JAVP: Joint-Aware Video Processing with Edge-Cloud Collaboration for DNN Inference
Zheming Yang, Wen Ji, Qi Guo, Zhi Wang
摘要
Currently, massive video inference tasks are processed through edge-cloud collaboration. However, the diverse scenarios make it difficult to allocate the inference tasks efficiently, resulting in many wasted resources. In this paper, we propose a joint-aware video processing (JAVP) architecture for edge-cloud collaboration. First, we develop a multiscale complexity-aware model for predicting task complexity and determining its suitability for edge or cloud servers. The task is subsequently efficiently scheduled to the appropriate servers by integrating complexity with an adaptive resource-aware optimization algorithm. For input tasks, JAVP can dynamically and intelligently select the most appropriate server. The evaluation results on public datasets show that JAVP can improve the through-put by more than 70% compared to traditional cloud-only solutions while meeting accuracy requirements. And JAVP can improve the accuracy by 3%-5% and reduce delay and energy consumption by 16%-50% compared to state-of-the-art edge-cloud solutions.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper10
- Server-Driven Video Streaming for Deep Learning InferenceKuntai Du, Ahsan Pervaiz, Xin Yuan, Aakanksha Chowdhery 等SIGCOMM 2020 · 被引用 238 次
- Joint Configuration Adaptation and Bandwidth Allocation for Edge-based Real-time Video AnalyticsCan Wang, Sheng Zhang, Yu Chen, Zhuzhong Qian 等INFOCOM 2020 · 被引用 223 次
- SurveilEdge: Real-time Video Query based on Collaborative Cloud-Edge Deep LearningShibo Wang, Shusen Yang, Cong ZhaoINFOCOM 2020 · 被引用 76 次
- CLIO: enabling automatic compilation of deep learning pipelines across IoT and cloudJin Huang, Colin Samplawski, Deepak Ganesan, Benjamin M. Marlin 等MobiCom 2020 · 被引用 72 次
- AdaMask: Enabling Machine-Centric Video Streaming with Adaptive Frame Masking for DNN Inference OffloadingShengzhong Liu, Tianshi Wang, Jinyang Li, Dachun Sun 等ACM MM 2022 · 被引用 45 次
相关 Paper
- AppealNet: An Efficient and Highly-Accurate Edge/Cloud Collaborative Architecture for DNN InferenceMin Li, Yu Li, Ye Tian, Li Jiang 等DAC 2021 · 被引用 37 次
- Elf: accelerate high-resolution mobile deep vision with content-aware parallel offloadingWuyang Zhang, Zhezhi He, Luyang Liu, Zhenhua Jia 等MobiCom 2021 · 被引用 171 次
- Optimal and Approximate Parallelism-Based Computation Offloading Algorithms for Real-Time Multimodal Learning at the EdgeQuan Chen, Ming Yi, Jing Li, Ning Li 等INFOCOM 2025 · 被引用 6 次
- FastVA: Deep Learning Video Analytics Through Edge Processing and NPU in MobileTianxiang Tan, Guohong CaoINFOCOM 2020 · 被引用 67 次
- HybridFlow: Resource-Adaptive Subtask Routing for Efficient Edge-Cloud LLM InferenceJiangwen Dong, Jiayu Li, Tianhang Zheng, Wanyu LINICML 2026 · 被引用 3 次
