SplitNets: Designing Neural Architectures for Efficient Distributed Computing on Head-Mounted Systems
Xin Dong, Barbara De Salvo, Meng Li, Chiao Liu, Zhongnan Qu, H. T. Kung, Ziyun Li
摘要
We design deep neural networks (DNNs) and corresponding networks' splittings to distribute DNNs' workload to camera sensors and a centralized aggregator on head mounted devices to meet system performance targets in inference accuracy and latency under the given hardware resource constraints. To achieve an optimal balance among computation, communication, and performance, a split-aware neural architecture search framework, SplitNets, is introduced to conduct model designing, splitting, and communication reduction simultaneously. We further extend the framework to multi-view systems for learning to fuse inputs from multiple camera sensors with optimal performance and systemic efficiency. We validate SplitNets for single-view system on ImageNet as well as multi-view system on 3D classification, and show that the SplitNets framework achieves state-of-the-art (SOTA) performance and system latency compared with existing approaches.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- CAMJ: Enabling System-Level Energy Modeling and Architectural Exploration for In-Sensor Visual ComputingTianrui Ma, Yu Feng, Xuan Zhang, Yuhao ZhuISCA 2023 · 被引用 17 次
- PixelRNN: In-pixel Recurrent Neural Networks for End-to-end-optimized Perception with Neural SensorsHaley M. So, Laurie Bose, Piotr Dudek, Gordon WetzsteinCVPR 2024 · 被引用 9 次
- On-Demand Container Partitioning for Distributed MLGiovanni Bartolomeo, Navidreza Asadi, Wolfgang Kellerer, Jörg Ott 等USENIX ATC 2025 · 被引用 3 次
它引用的顶会 Paper10
- Searching for MobileNetV3Andrew Howard, Ruoming Pang, Hartwig Adam, Quoc V. Le 等ICCV 2019 · 被引用 9,163 次
- Once-for-All: Train One Network and Specialize it for Efficient DeploymentHan Cai, Chuang Gan, Tianzhe Wang, Zhekai Zhang 等ICLR 2020 · 被引用 1,522 次
- Learned Step Size quantizationSteven K. Esser, Jeffrey L. McKinstry, Deepika Bablani, Rathinakumar Appuswamy 等ICLR 2020 · 被引用 1,037 次
- FairNAS: Rethinking Evaluation Fairness of Weight Sharing Neural Architecture SearchXiangxiang Chu, Bo Zhang, Ruijun XuICCV 2021 · 被引用 362 次
- Additive Powers-of-Two Quantization: An Efficient Non-uniform Discretization for Neural NetworksYuhang Li, Xin Dong, Wei WangICLR 2020 · 被引用 315 次
相关 Paper
- MTL-Split: Multi-Task Learning for Edge Devices using Split ComputingLuigi Capogrosso, Enrico Fraccaroli, Samarjit Chakraborty, Franco Fummi 等DAC 2024 · 被引用 12 次
- You only search once: on lightweight differentiable architecture search for resource-constrained embedded platformsXiangzhong Luo, Di Liu, Hao Kong, Shuo Huai 等DAC 2022 · 被引用 13 次
- H2H: heterogeneous model to heterogeneous system mapping with computation and communication awarenessXinyi Zhang, Cong Hao, Peipei Zhou, Alex K. Jones 等DAC 2022 · 被引用 21 次
- MARS: Exploiting Multi-Level Parallelism for DNN Workloads on Adaptive Multi-Accelerator SystemsGuan Shen, Jieru Zhao, Zeke Wang, Zhe Lin 等DAC 2023 · 被引用 5 次
- AutoShrink: A Topology-Aware NAS for Discovering Efficient Neural ArchitectureTunhou Zhang, Hsin-Pai Cheng, Zhenwen Li, Feng Yan 等AAAI 2020 · 被引用 9 次
