AccuMO: Accuracy-Centric Multitask Offloading in Edge-Assisted Mobile Augmented Reality
Z. Jonny Kong, Qiang Xu, Jiayi Meng, Y. Charlie Hu
Abstract
Immersive applications such as Augmented Reality (AR) and Mixed Reality (MR) often need to perform multiple latency-critical tasks on every frame captured by the camera, which all require results to be available within the current frame interval. While such tasks are increasingly supported by Deep Neural Networks (DNNs) offloaded to edge servers due to their high accuracy but heavy computation, prior work has largely focused on offloading one task at a time. Compared to offloading a single task, where more frequent offloading directly translates into higher task accuracy, offloading of multiple tasks competes for shared edge server resources, and hence faces the additional challenge of balancing the offloading frequencies of different tasks to maximize the overall accuracy and hence app QoE.
In this paper, we formulate this accuracy-centric multitask offloading problem, and present a framework that dynamically schedules the offloading of multiple DNN tasks from a mobile device to an edge server while optimizing the overall accuracy across tasks. Our design employs two novel ideas: (1) task-specific lightweight models that predict offloading accuracy drop as a function of offloading frequency and frame content, and (2) a general two-level control feedback loop that concurrently balances offloading among tasks and adapts between offloading and using local algorithms for each task. Evaluation results show that our framework improves the overall accuracy significantly in jointly offloading two core tasks in AR -depth estimation and odometry -by on average 7.6%-14.3% over the best baselines under different accuracy weight ratios.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 88bb04eb-03da-4cc2-8808-93f0fe348c64Cited by top-tier papers2
- ABO: Abandon Bayer Filter for Adaptive Edge Offloading in Responsive Augmented RealityYongxuan Han, Shengzhong Liu, Fan Wu, Guihai ChenWWW 2025 · 1 citation
- UrgenGo: Urgency-Aware Transparent GPU Kernel Launching for Autonomous DrivingHanqi Zhu, Wuyang Zhang, Xinran Zhang, Ziyang Tao et al.MobiCom 2025
Builds on17
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar et al.NeurIPS 2021 · 9,661 citations
- Serving DNNs like Clockwork: Performance Predictability from the Bottom UpArpan Gujarati, Reza Karimi, Safya Alzayat, Wei Hao et al.OSDI 2020 · 392 citations
- SPINN: synergistic progressive inference of neural networks over device and cloudStefanos Laskaridis, Stylianos I. Venieris, Mário Almeida, Ilias Leontiadis et al.MobiCom 2020 · 312 citations
- Reducto: On-Camera Filtering for Resource-Efficient Real-Time Video AnalyticsYuanqi Li, Arthi Padmanabhan, Pengzhan Zhao, Yufei Wang et al.SIGCOMM 2020 · 264 citations
- Server-Driven Video Streaming for Deep Learning InferenceKuntai Du, Ahsan Pervaiz, Xin Yuan, Aakanksha Chowdhery et al.SIGCOMM 2020 · 238 citations
Related papers
- AoDNN: An Auto-Offloading Approach to Optimize Deep Inference for Fostering Mobile WebYakun Huang, Xiuquan Qiao, Schahram Dustdar, Yan LiINFOCOM 2022 · 17 citations
- SEAR: Scaling Experiences in Multi-user Augmented RealityWenxiao Zhang, Bo Han, Pan HuiIEEE VR 2022 · 42 citations
- RTInfer: Real-Time Inference of Multiple DNNs on Edge GPUsRenjie Li, Tong Sun, Yi Gao, Wei DongICML 2026
- MTL-Split: Multi-Task Learning for Edge Devices using Split ComputingLuigi Capogrosso, Enrico Fraccaroli, Samarjit Chakraborty, Franco Fummi et al.DAC 2024 · 12 citations
- User Preference Based Energy-Aware Mobile AR System with Edge ComputingHaoxin Wang, Jiang (Linda) XieINFOCOM 2020 · 53 citations
