Stronger NAS with Weaker Predictors
Junru Wu, Xiyang Dai, Dongdong Chen, Yinpeng Chen, Mengchen Liu, Ye Yu, Zhangyang Wang, Zicheng Liu, Mei Chen, Lu Yuan
摘要
Neural Architecture Search (NAS) often trains and evaluates a large number of architectures. Recent predictor-based NAS approaches attempt to alleviate such heavy computation costs with two key steps: sampling some architecture-performance pairs and fitting a proxy accuracy predictor. Given limited samples, these predictors, however, are far from accurate to locate top architectures due to the difficulty of fitting the huge search space. This paper reflects on a simple yet crucial question: if our final goal is to find the best architecture, do we really need to model the whole space well?. We propose a paradigm shift from fitting the whole architecture space using one strong predictor, to progressively fitting a search path towards the high-performance sub-space through a set of weaker predictors. As a key property of the weak predictors, their probabilities of sampling better architectures keep increasing. Hence we only sample a few well-performed architectures guided by the previously learned predictor and estimate a new better weak predictor. This embarrassingly easy framework, dubbed WeakNAS, produces coarse-to-fine iteration to gradually refine the ranking of sampling space. Extensive experiments demonstrate that WeakNAS costs fewer samples to find top-performance architectures on NAS-Bench-101 and NAS-Bench-201. Compared to state-of-the-art (SOTA) predictor-based NAS methods, WeakNAS outperforms all with notable margins, e.g., requiring at least 7.5x less samples to find global optimal on NAS-Bench-101. WeakNAS can also absorb their ideas to boost performance more. Further, Weak-NAS strikes the new SOTA result of 81.3% in the ImageNet MobileNet Search Space. The code is available at: https://github.com/VITA-Group/WeakNAS . Recently, predictor-based NAS methods alleviate this problem with two key steps: one sampling step to sample some architecture-performance pairs, and another performance modeling step to fit the performance distribution by training a proxy accuracy predictor. An in-depth analysis of existing methods [2] found that most of those methods [5, 6, 17, [7] [8] [9] 18 ] consider these two steps independently and attempt to model the performance distribution over the whole architec-35th Conference on Neural Information Processing Systems (NeurIPS 2021).
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper14
- PINAT: A Permutation INvariance Augmented Transformer for NAS PredictorShun Lu, Yu Hu, Peihao Wang, Yan Han 等AAAI 2023 · 被引用 31 次
- Design Principle Transfer in Neural Architecture Search via Large Language ModelsXun Zhou, Xingyu Wu, Liang Feng, Zhichao Lu 等AAAI 2025 · 被引用 24 次
- Arch-Graph: Acyclic Architecture Relation Predictor for Task-Transferable Neural Architecture SearchMinbin Huang, Zhijian Huang, Changlin Li, Xin Chen 等CVPR 2022 · 被引用 20 次
- Bridge the Gap Between Architecture Spaces via A Cross-Domain PredictorYuqiao Liu, Yehui Tang, Zeqiong Lv, Yunhe Wang 等NeurIPS 2022 · 被引用 14 次
- Visual Analysis of Neural Architecture Spaces for Summarizing Design PrinciplesJun Yuan, Mengchen Liu, Fengyuan Tian, Shixia LiuIEEE VIS 2022 · 被引用 10 次
它引用的顶会 Paper14
- Searching for MobileNetV3Andrew Howard, Ruoming Pang, Hartwig Adam, Quoc V. Le 等ICCV 2019 · 被引用 9,163 次
- Once-for-All: Train One Network and Specialize it for Efficient DeploymentHan Cai, Chuang Gan, Tianzhe Wang, Zhekai Zhang 等ICLR 2020 · 被引用 1,522 次
- NAS-Bench-201: Extending the Scope of Reproducible Neural Architecture SearchXuanyi Dong, Yi YangICLR 2020 · 被引用 825 次
- Progressive Differentiable Architecture Search: Bridging the Depth Gap Between Search and EvaluationXin Chen, Lingxi Xie, Jun Wu, Qi TianICCV 2019 · 被引用 725 次
- BANANAS: Bayesian Optimization with Neural Architectures for Neural Architecture SearchColin White, Willie Neiswanger, Yash SavaniAAAI 2021 · 被引用 401 次
相关 Paper
- ReNAS: Relativistic Evaluation of Neural Architecture SearchYixing Xu, Yunhe Wang, Kai Han, Yehui Tang 等CVPR 2021
- Dynamic Ensemble of Low-Fidelity Experts: Mitigating NAS "Cold-Start"Junbo Zhao, Xuefei Ning, Enshu Liu, Binxin Ru 等AAAI 2023 · 被引用 5 次
- Generalized Global Ranking-Aware Neural Architecture Ranker for Efficient Image Classifier SearchBicheng Guo, Tao Chen, Shibo He, Haoyu Liu 等ACM MM 2022 · 被引用 21 次
- RANK-NOSH: Efficient Predictor-Based Architecture Search via Non-Uniform Successive HalvingRuochen Wang, Xiangning Chen, Minhao Cheng, Xiaocheng Tang 等ICCV 2021 · 被引用 14 次
- HyperNAS: Enhancing Architecture Representation for NAS Predictor via HypernetworkJindi Lv, Yuhao Zhou, Yuxin Tian, Qing Ye 等CVPR 2026 · 被引用 1 次
