FBNetV3: Joint Architecture-Recipe Search Using Predictor Pretraining
Xiaoliang Dai, Alvin Wan, Peizhao Zhang, Bichen Wu, Zijian He, Zhen Wei, Kan Chen, Yuandong Tian, Matthew Yu, Peter Vajda, Joseph E. Gonzalez
摘要
Neural Architecture Search (NAS) yields state-of-theart neural networks that outperform their best manuallydesigned counterparts. However, previous NAS methods search for architectures under one set of training hyperparameters (i.e., a training recipe), overlooking superior architecture-recipe combinations. To address this, we present Neural Architecture-Recipe Search (NARS) to search both (a) architectures and (b) their corresponding training recipes, simultaneously. NARS utilizes an accuracy predictor that scores architecture and training recipes jointly, guiding both sample selection and ranking. Furthermore, to compensate for the enlarged search space, we leverage "free" architecture statistics (e.g., FLOP count) to pretrain the predictor, significantly improving its sample efficiency and prediction reliability. After training the predictor via constrained iterative optimization, we run fast evolutionary searches in just CPU minutes to generate architecturerecipe pairs for a variety of resource constraints, called FBNetV3. FBNetV3 makes up a family of state-of-the-art compact neural networks that outperform both automatically and manually-designed competitors. For example, FB-NetV3 matches both EfficientNet and ResNeSt accuracy on ImageNet with up to 2.0× and 7.1× fewer FLOPs, respectively. Furthermore, FBNetV3 yields significant performance gains for downstream object detection tasks, improving mAP despite 18% fewer FLOPs and 34% fewer parameters than EfficientNet-based equivalents.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper15
- Rethinking Vision Transformers for MobileNet Size and SpeedYanyu Li, Ju Hu, Yang Wen, Georgios Evangelidis 等ICCV 2023 · 被引用 300 次
- Non-Probability Sampling Network for Stochastic Human Trajectory PredictionInhwan Bae, Jin-Hwi Park, Hae-Gon JeonCVPR 2022 · 被引用 69 次
- Rethinking Bias Mitigation: Fairer Architectures Make for Fairer Face RecognitionSamuel Dooley, Rhea Sanjay Sukthanker, John P. Dickerson, Colin White 等NeurIPS 2023 · 被引用 41 次
- Efficient Modulation for Vision NetworksXu Ma, Xiyang Dai, Jianwei Yang, Bin Xiao 等ICLR 2024 · 被引用 30 次
- Reinforcement Learning with Automated Auxiliary Loss SearchTairan He, Yuge Zhang, Kan Ren, Minghuan Liu 等NeurIPS 2022 · 被引用 22 次
它引用的顶会 Paper11
- Searching for MobileNetV3Andrew Howard, Ruoming Pang, Hartwig Adam, Quoc V. Le 等ICCV 2019 · 被引用 9,163 次
- Once-for-All: Train One Network and Specialize it for Efficient DeploymentHan Cai, Chuang Gan, Tianzhe Wang, Zhekai Zhang 等ICLR 2020 · 被引用 1,522 次
- AtomNAS: Fine-Grained End-to-End Neural Architecture SearchJieru Mei, Yingwei Li, Xiaochen Lian, Xiaojie Jin 等ICLR 2020 · 被引用 110 次
- Efficient Segmentation: Learning Downsampling Near Semantic BoundariesDmitrii Marin, Zijian He, Peter Vajda, Priyam Chatterjee 等ICCV 2019 · 被引用 107 次
- Neural Architecture Search Using Deep Neural Networks and Monte Carlo Tree SearchLinnan Wang, Yiyang Zhao, Yuu Jinnai, Yuandong Tian 等AAAI 2020 · 被引用 56 次
相关 Paper
- FBNetV2: Differentiable Neural Architecture Search for Spatial and Channel DimensionsAlvin Wan, Xiaoliang Dai, Peizhao Zhang, Zijian He 等CVPR 2020
- FP-NAS: Fast Probabilistic Neural Architecture SearchZhicheng Yan, Xiaoliang Dai, Peizhao Zhang, Yuandong Tian 等CVPR 2021
- AttentiveNAS: Improving Neural Architecture Search via Attentive SamplingDilin Wang, Meng Li, Chengyue Gong, Vikas ChandraCVPR 2021
- Fast and Practical Neural Architecture SearchJiequan Cui, Pengguang Chen, Ruiyu Li, Shu Liu 等ICCV 2019 · 被引用 69 次
- BN-NAS: Neural Architecture Search with Batch NormalizationBoyu Chen, Peixia Li, Baopu Li, Chen Lin 等ICCV 2021 · 被引用 35 次
