Distribution Consistent Neural Architecture Search
Junyi Pan, Chong Sun, Yizhou Zhou, Ying Zhang, Chen Li
摘要
Recent progress on neural architecture search (NAS) has demonstrated exciting results on automating deep network architecture designs. In order to overcome the unaffordable complexity of training each candidate architecture from scratch, the state-of-the-art one-shot NAS approaches adopt a weight-sharing strategy to improve training efficiency. Although the computational cost is greatly reduced, such oneshot process introduces a severe weight coupling problem that largely degrades the evaluation accuracy of each candidate. The existing approaches often address the problem by shrinking the search space, model distillation, or fewshot training. Instead, in this paper, we propose a novel distribution consistent one-shot neural architecture search algorithm. We first theoretically investigate how the weight coupling problem affects the network searching performance from a parameter distribution perspective, and then propose a novel supernet training strategy with a Distribution Consistent Constraint that can provide a good measurement for the extent to which two architectures can share weights. Our strategy optimizes the supernet through iteratively inferring network weights and corresponding local sharing states. Such joint optimization of supernet's weights and topologies can diminish the discrepancy between the weights inherited from the supernet and the ones that are trained with a stand-alone model. As a result, it enables a more accurate model evaluation phase and leads to a better searching performance. We conduct extensive experiments on benchmark datasets with multiple searching spaces. The resulting architecture achieves superior performance over the current state-of-the-art NAS algorithms with comparable search costs, which demonstrates the efficacy of our approach.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- DCLP: Neural Architecture Predictor with Curriculum Contrastive LearningShenghe Zheng, Hongzhi Wang, Tianyu MuAAAI 2024 · 被引用 7 次
- Moss: Proxy Model-based Full-Weight Aggregation in Federated Learning with Heterogeneous ModelsYifeng Cai, Ziqi Zhang, Ding Li, Yao Guo 等UbiComp 2025
它引用的顶会 Paper18
- Searching for MobileNetV3Andrew Howard, Ruoming Pang, Hartwig Adam, Quoc V. Le 等ICCV 2019 · 被引用 9,163 次
- Once-for-All: Train One Network and Specialize it for Efficient DeploymentHan Cai, Chuang Gan, Tianzhe Wang, Zhekai Zhang 等ICLR 2020 · 被引用 1,522 次
- NAS-Bench-201: Extending the Scope of Reproducible Neural Architecture SearchXuanyi Dong, Yi YangICLR 2020 · 被引用 825 次
- FairNAS: Rethinking Evaluation Fairness of Weight Sharing Neural Architecture SearchXiangxiang Chu, Bo Zhang, Ruijun XuICCV 2021 · 被引用 362 次
- One-Shot Neural Architecture Search via Self-Evaluated Template NetworkXuanyi Dong, Yi YangICCV 2019 · 被引用 206 次
相关 Paper
- Few-Shot Neural Architecture SearchYiyang Zhao, Linnan Wang, Yuandong Tian, Rodrigo Fonseca 等ICML 2021 · 被引用 100 次
- PA&DA: Jointly Sampling PAth and DAta for Consistent NASShun Lu, Yu Hu, Longxing Yang, Zihao Sun 等CVPR 2023
- SUMNAS: Supernet with Unbiased Meta-Features for Neural Architecture SearchHyeonmin Ha, Ji-Hoon Kim, Semin Park, Byung-Gon ChunICLR 2022 · 被引用 5 次
- HEP-NAS: Towards Efficient Few-shot Neural Architecture Search via Hierarchical Edge PartitioningJianfeng Li, Jiawen Zhang, Feng Wang, Lianbo MaAAAI 2025
- Overcoming Multi-Model Forgetting in One-Shot NAS With Diversity MaximizationMiao Zhang, Huiqi Li, Shirui Pan, Xiaojun Chang 等CVPR 2020
