Evaluating Efficient Performance Estimators of Neural Architectures
Xuefei Ning, Changcheng Tang, Wenshuo Li, Zixuan Zhou, Shuang Liang, Huazhong Yang, Yu Wang
摘要
Conducting efficient performance estimations of neural architectures is a major challenge in neural architecture search (NAS). To reduce the architecture training costs in NAS, one-shot estimators (OSEs) amortize the architecture training costs by sharing the parameters of one "supernet" between all architectures. Recently, zero-shot estimators (ZSEs) that involve no training are proposed to further reduce the architecture evaluation cost. Despite the high efficiency of these estimators, the quality of such estimations has not been thoroughly studied. In this paper, we conduct an extensive and organized assessment of OSEs and ZSEs on five NAS benchmarks: NAS-Bench-101/201/301, and NDS ResNet/ResNeXt-A. Specifically, we employ a set of NAS-oriented criteria to study the behavior of OSEs and ZSEs and reveal that they have certain biases and variances. After analyzing how and why the OSE estimations are unsatisfying, we explore how to mitigate the correlation gap of OSEs from several perspectives. Through our analysis, we give out suggestions for future application and development of efficient architecture performance estimators. Furthermore, the analysis framework proposed in our work could be utilized in future research to give a more comprehensive understanding of newly designed architecture performance estimators. All codes are available at https://github.com/walkerning/aw_nas [30] . How well the one-shot estimations are correlated with the standalone architecture performances is essential for the efficacy of NAS methods. Despite the widespread use of OSEs, studies [50] have 35th Conference on Neural Information Processing Systems (NeurIPS 2021),
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper28
- ProxyBO: Accelerating Neural Architecture Search via Bayesian Optimization with Zero-Cost ProxiesYu Shen, Yang Li, Jian Zheng, Wentao Zhang 等AAAI 2023 · 被引用 43 次
- AutoCTS+: Joint Neural Architecture and Hyperparameter Search for Correlated Time Series ForecastingXinle Wu, Dalin Zhang, Miao Zhang, Chenjuan Guo 等SIGMOD 2023 · 被引用 36 次
- PINAT: A Permutation INvariance Augmented Transformer for NAS PredictorShun Lu, Yu Hu, Peihao Wang, Yan Han 等AAAI 2023 · 被引用 31 次
- Analyzing and Mitigating Interference in Neural Architecture SearchJin Xu, Xu Tan, Kaitao Song, Renqian Luo 等ICML 2022 · 被引用 30 次
- LiteTransformerSearch: Training-free Neural Architecture Search for Efficient Language ModelsMojan Javaheripi, Gustavo de Rosa, Subhabrata Mukherjee, Shital Shah 等NeurIPS 2022 · 被引用 27 次
它引用的顶会 Paper17
- Pruning neural networks without any data by iteratively conserving synaptic flowHidenori Tanaka, Daniel Kunin, Daniel L. K. Yamins, Surya GanguliNeurIPS 2020 · 被引用 884 次
- NAS-Bench-201: Extending the Scope of Reproducible Neural Architecture SearchXuanyi Dong, Yi YangICLR 2020 · 被引用 825 次
- Picking Winning Tickets Before Training by Preserving Gradient FlowChaoqi Wang, Guodong Zhang, Roger B. GrosseICLR 2020 · 被引用 743 次
- Progressive Differentiable Architecture Search: Bridging the Depth Gap Between Search and EvaluationXin Chen, Lingxi Xie, Jun Wu, Qi TianICCV 2019 · 被引用 725 次
- Neural Architecture Search without TrainingJoe Mellor, Jack Turner, Amos Storkey, Elliot J. CrowleyICML 2021 · 被引用 477 次
相关 Paper
- Few-Shot Neural Architecture SearchYiyang Zhao, Linnan Wang, Yuandong Tian, Rodrigo Fonseca 等ICML 2021 · 被引用 100 次
- Distribution Consistent Neural Architecture SearchJunyi Pan, Chong Sun, Yizhou Zhou, Ying Zhang 等CVPR 2022 · 被引用 9 次
- GreedyNAS: Towards Fast One-Shot NAS With Greedy SupernetShan You, Tao Huang, Mingmin Yang, Fei Wang 等CVPR 2020
- ZiCo: Zero-shot NAS via inverse Coefficient of Variation on GradientsGuihong Li, Yuedong Yang, Kartikeya Bhardwaj, Radu MarculescuICLR 2023 · 被引用 19 次
- One-Shot Neural Architecture Search via Self-Evaluated Template NetworkXuanyi Dong, Yi YangICCV 2019 · 被引用 206 次
