Flexible high-resolution object detection on edge devices with tunable latency
Shiqi Jiang, Zhiqi Lin, Yuanchun Li, Yuanchao Shu, Yunxin Liu
摘要
Object detection is a fundamental building block of video analytics applications. While Neural Networks (NNs)-based object detection models have shown excellent accuracy on benchmark datasets, they are not well positioned for high-resolution images inference on resource-constrained edge devices. Common approaches, including down-sampling inputs and scaling up neural networks, fall short of adapting to video content changes and various latency requirements. This paper presents Remix, a flexible framework for high-resolution object detection on edge devices. Remix takes as input a latency budget, and come up with an image partition and model execution plan which runs off-the-shelf neural networks on non-uniformly partitioned image blocks. As a result, it maximizes the overall detection accuracy by allocating various amount of compute power onto different areas of an image. We evaluate Remix on public dataset as well as real-world videos collected by ourselves. Experimental results show that Remix can either improve the detection accuracy by 18%-120% for a given latency budget, or achieve up to 8.1× inference speedup with accuracy on par with the state-of-the-art NNs.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper17
- Mobile Foundation Model as FirmwareJinliang Yuan, Chen Yang, Dongqi Cai, Shihe Wang 等MobiCom 2024 · 被引用 40 次
- Cross-Camera Inference on the Constrained EdgeJingzong Li, Libin Liu, Hong Xu, Shudeng Wu 等INFOCOM 2023 · 被引用 31 次
- DeepPerform: An Efficient Approach for Performance Testing of Resource-Constrained Neural NetworksSimin Chen, Mirazul Haque, Cong Liu, Wei YangASE 2022 · 被引用 19 次
- Vulcan: Automatic Query Planning for Live ML AnalyticsYiwen Zhang, Xumiao Zhang, Ganesh Ananthanarayanan, Anand P. Iyer 等NSDI 2024 · 被引用 17 次
- SwapMoE: Serving Off-the-shelf MoE-based Large Language Models with Tunable Memory BudgetRui Kong, Yuanchun Li, Qingtian Feng, Weijun Wang 等ACL 2024 · 被引用 12 次
它引用的顶会 Paper12
- Once-for-All: Train One Network and Specialize it for Efficient DeploymentHan Cai, Chuang Gan, Tianzhe Wang, Zhekai Zhang 等ICLR 2020 · 被引用 1,522 次
- Elf: accelerate high-resolution mobile deep vision with content-aware parallel offloadingWuyang Zhang, Zhezhi He, Luyang Liu, Zhenhua Jia 等MobiCom 2021 · 被引用 171 次
- Heimdall: mobile GPU coordination platform for augmented reality applicationsJuheon Yi, Youngki LeeMobiCom 2020 · 被引用 71 次
- EagleEye: wearable camera-based person identification in crowded urban spacesJuheon Yi, Sunghyun Choi, Youngki LeeMobiCom 2020 · 被引用 69 次
- Mistify: Automating DNN Model Porting for On-Device Inference at the EdgePeizhen Guo, Bo Hu, Wenjun HuNSDI 2021 · 被引用 69 次
相关 Paper
- LiteReconfig: cost and content aware reconfiguration of video object detection systems for mobile GPUsRan Xu, Jayoung Lee, Pengcheng Wang, Saurabh Bagchi 等EuroSys 2022 · 被引用 24 次
- ResMap: Exploiting Sparse Residual Feature Map for Accelerating Cross-Edge Video AnalyticsNing Chen, Shuai Zhang, Sheng Zhang, Yuting Yan 等INFOCOM 2023 · 被引用 11 次
- Edge-assisted Online On-device Object Detection for Real-time Video AnalyticsMengxi Hanyao, Yibo Jin, Zhuzhong Qian, Sheng Zhang 等INFOCOM 2021 · 被引用 82 次
- FlexPatch: Fast and Accurate Object Detection for On-device High-Resolution Live Video AnalyticsKichang Yang, Juheon Yi, Kyungjin Lee, Youngki LeeINFOCOM 2022 · 被引用 43 次
- RECL: Responsive Resource-Efficient Continuous Learning for Video AnalyticsMehrdad Khani Shirkoohi, Ganesh Ananthanarayanan, Kevin Hsieh, Junchen Jiang 等NSDI 2023
