AsyMo: scalable and efficient deep-learning inference on asymmetric mobile CPUs
Manni Wang, Shaohua Ding, Ting Cao, Yunxin Liu, Fengyuan Xu
2021年份
69被引次数
17顶会引用
摘要
On-device deep learning (DL) inference has attracted vast interest. Mobile CPUs are the most common hardware for on-device inference and many inference frameworks have been developed for them. Yet, due to the hardware complexity, DL inference on mobile CPUs suffers from two common issues: the poor performance scalability on the asymmetric multiprocessor, and energy inefficiency.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper17
- Flexible high-resolution object detection on edge devices with tunable latencyShiqi Jiang, Zhiqi Lin, Yuanchun Li, Yuanchao Shu 等MobiCom 2021 · 被引用 103 次
- A Comprehensive Benchmark of Deep Learning Libraries on Mobile DevicesQiyang Zhang, Xiang Li, Xiangying Che, Xiao Ma 等WWW 2022 · 被引用 61 次
- AdaptiveNet: Post-deployment Neural Architecture Adaptation for Diverse Edge EnvironmentsHao Wen, Yuanchun Li, Zunshuai Zhang, Shiqi Jiang 等MobiCom 2023 · 被引用 55 次
- A Workload-Aware DVFS Robust to Concurrent Tasks for Mobile DevicesChengdong Lin, Kun Wang, Zhenjiang Li, Yu PuMobiCom 2023 · 被引用 52 次
- Galaxy: A Resource-Efficient Collaborative Edge AI System for In-situ Transformer InferenceShengyuan Ye, Jiangsu Du, Liekang Zeng, Wenzhong Ou 等INFOCOM 2024 · 被引用 43 次
相关 Paper
- InfScaler: Enabling Efficient ML Inference Serving on Multi-Accelerator Edge Devices via Asymmetric Auto-ScalingBorui Li, Tiange Xia, Shuai Wang, Shuai WangDAC 2025 · 被引用 2 次
- Unleash All Cores: Asymmetry-Aware Scalable DNN Inference on Mobile CPUsQianlong Sang, Puyi He, Huanghuang Liang, Yili Gong 等OSDI 2026
- AutoScale: Energy Efficiency Optimization for Stochastic Edge Inference Using Reinforcement LearningYoung Geun Kim, Carole-Jean WuMICRO 2020 · 被引用 80 次
- Mosaic: Exploiting Instruction-Level Parallelism on Deep Learning Accelerators with iTex TessellationJianxing Xu, Yuanbo Wen, Zikang Liu, Ruibai Xu 等ASPLOS 2025 · 被引用 2 次
- Minimizing Latency for Multi-DNN Inference on Resource-Limited CPU-Only Edge DevicesTao Wang, Tuo Shi, Xiulong Liu, Jianping Wang 等INFOCOM 2024 · 被引用 9 次
