Boosting Order-Preserving and Transferability for Neural Architecture Search: A Joint Architecture Refined Search and Fine-Tuning Approach
Beichen Zhang, Xiaoxing Wang, Xiaohan Qin, Junchi Yan
摘要
Supernet is a core component in many recent Neural Architecture Search (NAS) methods. It not only helps embody the search space but also provides a (relative) estimation of the final performance of candidate architectures. Thus, it is critical that the top architectures ranked by a supernet should be consistent with those ranked by true performance, which is known as the order-preserving ability. In this work, we analyze the order-preserving ability on the whole search space (global) and a sub-space of top architectures (local), and empirically show that the local order-preserving for current two-stage NAS methods still need to be improved. To rectify this, we propose a novel concept of Supernet Shifting, a refined search strategy combining architecture searching with supernet fine-tuning. Specifically, apart from evaluating, the training loss is also accumulated in searching and the supernet is updated every iteration. Since superior architectures are sampled more frequently in evolutionary searching, the supernet is encouraged to focus on top architectures, thus improving local order-preserving. Besides, a pre-trained supernet is often un-reusable for one-shot methods. We show that Supernet Shifting can fulfill transferring supernet to a new dataset. Specifically, the last classifier layer will be unset and trained through evolutionary searching. Comprehensive experiments show that our method has better order-preserving ability and can find a dominating architecture. Moreover, the pre-trained supernet can be easily transferred into a new dataset with no loss of performance.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- ReLIZO: Sample Reusable Linear Interpolation-based Zeroth-order OptimizationXiaoxing Wang, Xiaohan Qin, Xiaokang Yang, Junchi YanNeurIPS 2024 · 被引用 10 次
- Per-Architecture Training-Free Metric Optimization for Neural Architecture SearchMingzhuo Lin, Jianping LuoNeurIPS 2025 · 被引用 3 次
- HEP-NAS: Towards Efficient Few-shot Neural Architecture Search via Hierarchical Edge PartitioningJianfeng Li, Jiawen Zhang, Feng Wang, Lianbo MaAAAI 2025
- Conical Visual Concentration for Efficient Large Vision-Language ModelsLong Xing, Qidong Huang, Xiaoyi Dong, Jiajie Lu 等CVPR 2025
它引用的顶会 Paper16
- Picking Winning Tickets Before Training by Preserving Gradient FlowChaoqi Wang, Guodong Zhang, Roger B. GrosseICLR 2020 · 被引用 743 次
- Progressive Differentiable Architecture Search: Bridging the Depth Gap Between Search and EvaluationXin Chen, Lingxi Xie, Jun Wu, Qi TianICCV 2019 · 被引用 725 次
- Neural Architecture Search without TrainingJoe Mellor, Jack Turner, Amos Storkey, Elliot J. CrowleyICML 2021 · 被引用 477 次
- FairNAS: Rethinking Evaluation Fairness of Weight Sharing Neural Architecture SearchXiangxiang Chu, Bo Zhang, Ruijun XuICCV 2021 · 被引用 362 次
- Few-Shot Neural Architecture SearchYiyang Zhao, Linnan Wang, Yuandong Tian, Rodrigo Fonseca 等ICML 2021 · 被引用 100 次
相关 Paper
- SUMNAS: Supernet with Unbiased Meta-Features for Neural Architecture SearchHyeonmin Ha, Ji-Hoon Kim, Semin Park, Byung-Gon ChunICLR 2022 · 被引用 5 次
- Distribution Consistent Neural Architecture SearchJunyi Pan, Chong Sun, Yizhou Zhou, Ying Zhang 等CVPR 2022 · 被引用 9 次
- Pi-NAS: Improving Neural Architecture Search by Reducing Supernet Training Consistency ShiftJiefeng Peng, Jiqi Zhang, Changlin Li, Guangrun Wang 等ICCV 2021 · 被引用 20 次
- Overcoming Multi-Model Forgetting in One-Shot NAS With Diversity MaximizationMiao Zhang, Huiqi Li, Shirui Pan, Xiaojun Chang 等CVPR 2020
- One-Shot Neural Ensemble Architecture Search by Diversity-Guided Search Space ShrinkingMinghao Chen, Jianlong Fu, Haibin LingCVPR 2021
