Genie in the Model: Automatic Generation of Human-in-the-Loop Deep Neural Networks for Mobile Applications
Yanfei Wang, Zhiwen Yu, Sicong Liu, Zimu Zhou, Bin Guo
摘要
Advances in deep neural networks (DNNs) have fostered a wide spectrum of intelligent mobile applications ranging from voice assistants on smartphones to augmented reality with smart-glasses. To deliver high-quality services, these DNNs should operate on resource-constrained mobile platforms and yield consistent performance in open environments. However, DNNs are notoriously resource-intensive, and often suffer from performance degradation in real-world deployments. Existing research strives to optimize the resource-performance trade-off of DNNs by compressing the model without notably compromising its inference accuracy. Accordingly, the accuracy of these compressed DNNs is bounded by the original ones, leading to more severe accuracy drop in challenging yet common scenarios such as low-resolution, small-size, and motion-blur. In this paper, we propose to push forward the frontiers of the DNN performance-resource trade-off by introducing human intelligence as a new design dimension. To this end, we explore human-in-the-loop DNNs (H-DNNs) and their automatic performance-resource optimization. We present H-Gen, an automatic H-DNN compression framework that incorporates human participation as a new hyperparameter for accurate and efficient DNN generation. It involves novel hyperparameter formulation, metric calculation, and search strategy in the context of automatic H-DNN generation. We also propose human participation mechanisms for three common DNN architectures to showcase the feasibility of H-Gen. Extensive experiments on twelve categories of challenging samples with three common DNN structures demonstrate the superiority of H-Gen in terms of the overall trade-off between performance (accuracy, latency), and resource (storage, energy, human labour).
CCS Concepts: • Human-centered computing → Ubiquitous and mobile computing systems and tools.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper14
- Active Learning for Deep Object Detection via Probabilistic ModelingJiwoong Choi, Ismail Elezi, Hyuk-Jae Lee, Clément Farabet 等ICCV 2021 · 被引用 144 次
- Deep Reinforcement Active Learning for Human-in-the-Loop Person Re-IdentificationZimo Liu, Jingya Wang, Shaogang Gong, Dacheng Tao 等ICCV 2019 · 被引用 117 次
- Attend and Discriminate: Beyond the State-of-the-Art for Human Activity Recognition Using Wearable SensorsAlireza Abedin, Mahsa Ehsanpour, Qinfeng Shi, Hamid Rezatofighi 等UbiComp 2021 · 被引用 104 次
- Multinomial Distribution Learning for Effective Neural Architecture SearchXiawu Zheng, Rongrong Ji, Lang Tang, Baochang Zhang 等ICCV 2019 · 被引用 100 次
- Human-in-the-loop Outlier DetectionChengliang Chai, Lei Cao, Guoliang Li, Jian Li 等SIGMOD 2020 · 被引用 57 次
相关 Paper
- Automatic Neural Network Compression by Sparsity-Quantization Joint Learning: A Constrained Optimization-Based ApproachHaichuan Yang, Shupeng Gui, Yuhao Zhu, Ji LiuCVPR 2020
- Auto Graph Encoder-Decoder for Neural Network PruningSixing Yu, Arya Mazaheri, Ali JannesariICCV 2021 · 被引用 47 次
- AdaSpring: Context-adaptive and Runtime-evolutionary Deep Model Compression for Mobile ApplicationsSicong Liu, Bin Guo, Ke Ma, Zhiwen Yu 等UbiComp 2021 · 被引用 30 次
- AoDNN: An Auto-Offloading Approach to Optimize Deep Inference for Fostering Mobile WebYakun Huang, Xiuquan Qiao, Schahram Dustdar, Yan LiINFOCOM 2022 · 被引用 17 次
- Scalable Super-Resolution Neural OperatorLei Han, Xuesong ZhangACM MM 2024 · 被引用 2 次
