AsyMo: scalable and efficient deep-learning inference on asymmetric mobile CPUs
Manni Wang, Shaohua Ding, Ting Cao, Yunxin Liu, Fengyuan Xu
2021Year
69Citations
17Top-tier citations
Abstract
On-device deep learning (DL) inference has attracted vast interest. Mobile CPUs are the most common hardware for on-device inference and many inference frameworks have been developed for them. Yet, due to the hardware complexity, DL inference on mobile CPUs suffers from two common issues: the poor performance scalability on the asymmetric multiprocessor, and energy inefficiency.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get e6b948ed-8b8d-493d-b155-66b40ccf6270Cited by top-tier papers17
- Flexible high-resolution object detection on edge devices with tunable latencyShiqi Jiang, Zhiqi Lin, Yuanchun Li, Yuanchao Shu et al.MobiCom 2021 · 103 citations
- A Comprehensive Benchmark of Deep Learning Libraries on Mobile DevicesQiyang Zhang, Xiang Li, Xiangying Che, Xiao Ma et al.WWW 2022 · 61 citations
- AdaptiveNet: Post-deployment Neural Architecture Adaptation for Diverse Edge EnvironmentsHao Wen, Yuanchun Li, Zunshuai Zhang, Shiqi Jiang et al.MobiCom 2023 · 55 citations
- A Workload-Aware DVFS Robust to Concurrent Tasks for Mobile DevicesChengdong Lin, Kun Wang, Zhenjiang Li, Yu PuMobiCom 2023 · 52 citations
- Galaxy: A Resource-Efficient Collaborative Edge AI System for In-situ Transformer InferenceShengyuan Ye, Jiangsu Du, Liekang Zeng, Wenzhong Ou et al.INFOCOM 2024 · 43 citations
Related papers
- InfScaler: Enabling Efficient ML Inference Serving on Multi-Accelerator Edge Devices via Asymmetric Auto-ScalingBorui Li, Tiange Xia, Shuai Wang, Shuai WangDAC 2025 · 2 citations
- Unleash All Cores: Asymmetry-Aware Scalable DNN Inference on Mobile CPUsQianlong Sang, Puyi He, Huanghuang Liang, Yili Gong et al.OSDI 2026
- AutoScale: Energy Efficiency Optimization for Stochastic Edge Inference Using Reinforcement LearningYoung Geun Kim, Carole-Jean WuMICRO 2020 · 80 citations
- Mosaic: Exploiting Instruction-Level Parallelism on Deep Learning Accelerators with iTex TessellationJianxing Xu, Yuanbo Wen, Zikang Liu, Ruibai Xu et al.ASPLOS 2025 · 2 citations
- Minimizing Latency for Multi-DNN Inference on Resource-Limited CPU-Only Edge DevicesTao Wang, Tuo Shi, Xiulong Liu, Jianping Wang et al.INFOCOM 2024 · 9 citations
