Hyperscale Hardware Optimized Neural Architecture Search
Sheng Li, Garrett Andersen, Tao Chen, Liqun Cheng, Julian Grady, Da Huang, Quoc V. Le, Andrew Li, Xin Li, Yang Li, Chen Liang, Yifeng Lu
摘要
Recent advances in machine learning have leveraged dramatic increases in computational power, a trend expected to continue in the future. This paper introduces the first Hyperscale Hardware Optimized Neural Architecture Search (H2O-NAS) to automatically design accurate and performant machine learning models tailored to the underlying hardware architecture. H2O-NAS consists of three key components: a new massively parallel “one-shot” search algorithm with intelligent weight sharing, which can scale to search spaces of O(10280) and handle large volumes of production traffic; hardware-optimized search spaces for diverse ML models on heterogeneous hardware; and a novel two-phase hybrid performance model and a multi-objective reward function optimized for large scale deployments.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper13
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Searching for MobileNetV3Andrew Howard, Ruoming Pang, Hartwig Adam, Quoc V. Le 等ICCV 2019 · 被引用 9,163 次
- GShard: Scaling Giant Models with Conditional Computation and Automatic ShardingDmitry Lepikhin, HyoukJoong Lee, Yuanzhong Xu, Dehao Chen 等ICLR 2021 · 被引用 1,954 次
- CoAtNet: Marrying Convolution and Attention for All Data SizesZihang Dai, Hanxiao Liu, Quoc V. Le, Mingxing TanNeurIPS 2021 · 被引用 1,747 次
- Once-for-All: Train One Network and Specialize it for Efficient DeploymentHan Cai, Chuang Gan, Tianzhe Wang, Zhekai Zhang 等ICLR 2020 · 被引用 1,522 次
相关 Paper
- Rapid Model Architecture Adaption for Meta-LearningYiren Zhao, Xitong Gao, Ilia Shumailov, Nicolò Fusi 等NeurIPS 2022 · 被引用 8 次
- NAS-SE: Designing A Highly-Efficient In-Situ Neural Architecture Search Engine for Large-Scale DeploymentQiyu Wan, Lening Wang, Jing Wang, Shuaiwen Leon Song 等MICRO 2023 · 被引用 2 次
- NASA: Accelerating Neural Network Design with a NAS ProcessorXiaohan Ma, Chang Si, Ying Wang, Cheng Liu 等ISCA 2021 · 被引用 8 次
- NAS-Bench-1Shot1: Benchmarking and Dissecting One-shot Neural Architecture SearchArber Zela, Julien Siems, Frank HutterICLR 2020 · 被引用 156 次
- PreNAS: Preferred One-Shot Learning Towards Efficient Neural Architecture SearchHaibin Wang, Ce Ge, Hesen Chen, Xiuyu SunICML 2023 · 被引用 27 次
