GreedyNAS: Towards Fast One-Shot NAS With Greedy Supernet
Shan You, Tao Huang, Mingmin Yang, Fei Wang, Chen Qian, Changshui Zhang
Abstract
Training a supernet matters for one-shot neural architecture search (NAS) methods since it serves as a basic performance estimator for different architectures (paths). Current methods mainly hold the assumption that a supernet should give a reasonable ranking over all paths. They thus treat all paths equally, and spare much effort to train paths. However, it is harsh for a single supernet to evaluate accurately on such a huge-scale search space (e.g., 7^21). In this paper, instead of covering all paths, we ease the burden of supernet by encouraging it to focus more on evaluation of those potentially-good ones, which are identified using a surrogate portion of validation data. Concretely, during training, we propose a multi-path sampling strategy with rejection, and greedily filter the weak paths. The training efficiency is thus boosted since the training space has been greedily shrunk from all paths to those potentially-good ones. Moreover, we further adopt an exploration and exploitation policy by introducing an empirical candidate path pool. Our proposed method GreedyNAS is easy-to-follow, and experimental results on ImageNet dataset indicate that it can achieve better Top-1 accuracy under same search space and FLOPs or latency level, but with only 60% of supernet training cost. By searching on a larger space, our GreedyNAS can also obtain new state-of-the-art architectures.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 19e708c6-307e-4d69-b1ba-9b0026d1e0eaCited by top-tier papers10
- Cream of the Crop: Distilling Prioritized Paths For One-Shot Neural Architecture SearchHouwen Peng, Hao Du, Hongyuan Yu, Qi Li et al.NeurIPS 2020 · 76 citations
- Data-Free Knowledge Distillation with Soft Targeted Transfer Set SynthesisZi WangAAAI 2021 · 35 citations
- Neural Architecture Search as Sparse SupernetYan Wu, Aoming Liu, Zhiwu Huang, Siwei Zhang et al.AAAI 2021 · 26 citations
- NASPipe: high performance and reproducible pipeline parallel supernet training via causal synchronous parallelismShixiong Zhao, Fanxin Li, Xusheng Chen, Tianxiang Shen et al.ASPLOS 2022 · 11 citations
- Distribution Consistent Neural Architecture SearchJunyi Pan, Chong Sun, Yizhou Zhou, Ying Zhang et al.CVPR 2022 · 9 citations
Builds on8
- FairNAS: Rethinking Evaluation Fairness of Weight Sharing Neural Architecture SearchXiangxiang Chu, Bo Zhang, Ruijun XuICCV 2021 · 362 citations
- Deep Comprehensive Correlation Mining for Image ClusteringJianlong Wu, Keyu Long, Fei Wang, Chen Qian et al.ICCV 2019 · 191 citations
- Multinomial Distribution Learning for Effective Neural Architecture SearchXiawu Zheng, Rongrong Ji, Lang Tang, Baochang Zhang et al.ICCV 2019 · 100 citations
- Reborn Filters: Pruning Convolutional Neural Networks with Limited DataYehui Tang, Shan You, Chang Xu, Jin Han et al.AAAI 2020 · 33 citations
- Learning Student Networks with Few DataShumin Kong, Tianyu Guo, Shan You, Chang XuAAAI 2020 · 12 citations
Related papers
- GreedyNASv2: Greedier Search with a Greedy Path FilterTao Huang, Shan You, Fei Wang, Chen Qian et al.CVPR 2022 · 16 citations
- Few-Shot Neural Architecture SearchYiyang Zhao, Linnan Wang, Yuandong Tian, Rodrigo Fonseca et al.ICML 2021 · 100 citations
- Searching by Generating: Flexible and Efficient One-Shot NAS With Architecture GeneratorSian-Yao Huang, Wei-Ta ChuCVPR 2021
- PA&DA: Jointly Sampling PAth and DAta for Consistent NASShun Lu, Yu Hu, Longxing Yang, Zihao Sun et al.CVPR 2023
- SUMNAS: Supernet with Unbiased Meta-Features for Neural Architecture SearchHyeonmin Ha, Ji-Hoon Kim, Semin Park, Byung-Gon ChunICLR 2022 · 5 citations
