GreedyNAS: Towards Fast One-Shot NAS With Greedy Supernet
Shan You, Tao Huang, Mingmin Yang, Fei Wang, Chen Qian, Changshui Zhang
摘要
Training a supernet matters for one-shot neural architecture search (NAS) methods since it serves as a basic performance estimator for different architectures (paths). Current methods mainly hold the assumption that a supernet should give a reasonable ranking over all paths. They thus treat all paths equally, and spare much effort to train paths. However, it is harsh for a single supernet to evaluate accurately on such a huge-scale search space (e.g., 7^21). In this paper, instead of covering all paths, we ease the burden of supernet by encouraging it to focus more on evaluation of those potentially-good ones, which are identified using a surrogate portion of validation data. Concretely, during training, we propose a multi-path sampling strategy with rejection, and greedily filter the weak paths. The training efficiency is thus boosted since the training space has been greedily shrunk from all paths to those potentially-good ones. Moreover, we further adopt an exploration and exploitation policy by introducing an empirical candidate path pool. Our proposed method GreedyNAS is easy-to-follow, and experimental results on ImageNet dataset indicate that it can achieve better Top-1 accuracy under same search space and FLOPs or latency level, but with only 60% of supernet training cost. By searching on a larger space, our GreedyNAS can also obtain new state-of-the-art architectures.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Cream of the Crop: Distilling Prioritized Paths For One-Shot Neural Architecture SearchHouwen Peng, Hao Du, Hongyuan Yu, Qi Li 等NeurIPS 2020 · 被引用 76 次
- Data-Free Knowledge Distillation with Soft Targeted Transfer Set SynthesisZi WangAAAI 2021 · 被引用 35 次
- Neural Architecture Search as Sparse SupernetYan Wu, Aoming Liu, Zhiwu Huang, Siwei Zhang 等AAAI 2021 · 被引用 26 次
- NASPipe: high performance and reproducible pipeline parallel supernet training via causal synchronous parallelismShixiong Zhao, Fanxin Li, Xusheng Chen, Tianxiang Shen 等ASPLOS 2022 · 被引用 11 次
- Distribution Consistent Neural Architecture SearchJunyi Pan, Chong Sun, Yizhou Zhou, Ying Zhang 等CVPR 2022 · 被引用 9 次
它引用的顶会 Paper8
- FairNAS: Rethinking Evaluation Fairness of Weight Sharing Neural Architecture SearchXiangxiang Chu, Bo Zhang, Ruijun XuICCV 2021 · 被引用 362 次
- Deep Comprehensive Correlation Mining for Image ClusteringJianlong Wu, Keyu Long, Fei Wang, Chen Qian 等ICCV 2019 · 被引用 191 次
- Multinomial Distribution Learning for Effective Neural Architecture SearchXiawu Zheng, Rongrong Ji, Lang Tang, Baochang Zhang 等ICCV 2019 · 被引用 100 次
- Reborn Filters: Pruning Convolutional Neural Networks with Limited DataYehui Tang, Shan You, Chang Xu, Jin Han 等AAAI 2020 · 被引用 33 次
- Learning Student Networks with Few DataShumin Kong, Tianyu Guo, Shan You, Chang XuAAAI 2020 · 被引用 12 次
相关 Paper
- GreedyNASv2: Greedier Search with a Greedy Path FilterTao Huang, Shan You, Fei Wang, Chen Qian 等CVPR 2022 · 被引用 16 次
- Few-Shot Neural Architecture SearchYiyang Zhao, Linnan Wang, Yuandong Tian, Rodrigo Fonseca 等ICML 2021 · 被引用 100 次
- Searching by Generating: Flexible and Efficient One-Shot NAS With Architecture GeneratorSian-Yao Huang, Wei-Ta ChuCVPR 2021
- PA&DA: Jointly Sampling PAth and DAta for Consistent NASShun Lu, Yu Hu, Longxing Yang, Zihao Sun 等CVPR 2023
- SUMNAS: Supernet with Unbiased Meta-Features for Neural Architecture SearchHyeonmin Ha, Ji-Hoon Kim, Semin Park, Byung-Gon ChunICLR 2022 · 被引用 5 次
