-DARTS: Mitigating Performance Collapse by Harmonizing Operation Selection among Cells
Sajad Movahedi, Melika Adabinejad, Ayyoob Imani, Arezou Keshavarz, Mostafa Dehghani, Azadeh Shakery, Babak Nadjar Araabi
摘要
Differentiable neural architecture search (DARTS) is a popular method for neural architecture search (NAS), which performs cell-search and utilizes continuous relaxation to improve the search efficiency via gradient-based optimization. The main shortcoming of DARTS is performance collapse, where the discovered architecture suffers from a pattern of declining quality during search. Performance collapse has become an important topic of research, with many methods trying to solve the issue through either regularization or fundamental changes to DARTS. However, the weight-sharing framework used for cell-search in DARTS and the convergence of architecture parameters has not been analyzed yet. In this paper, we provide a thorough and novel theoretical and empirical analysis on DARTS and its point of convergence. We show that DARTS suffers from a specific structural flaw due to its weight-sharing framework that limits the convergence of DARTS to saturation points of the softmax function. This point of convergence gives an unfair advantage to layers closer to the output in choosing the optimal architecture, causing performance collapse. We then propose two new regularization terms that aim to prevent performance collapse by harmonizing operation selection via aligning gradients of layers. Experimental results on six different search spaces and three different datasets show that our method (Λ-DARTS) does indeed prevent performance collapse, providing justification for our theoretical analysis and the proposed remedy. We have published our code at https://github.com/dr-faustus/Lambda-DARTS . RELATED WORK DARTS (Liu et al., 2019) proposed a continuous and differentiable search space through weighting a fixed set of operations to make NAS more scalable. It trains a super-graph with gradient descent and chooses the sub-graph consisted of weightiest operation edges. Its simplicity made it very popular and many variations emerged to address its theoretical and empirical setbacks:
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- HyperNAS: Enhancing Architecture Representation for NAS Predictor via HypernetworkJindi Lv, Yuhao Zhou, Yuxin Tian, Qing Ye 等CVPR 2026 · 被引用 1 次
- NADER: Neural Architecture Design via Multi-Agent CollaborationZekang Yang, Wang Zeng, Sheng Jin, Chen Qian 等CVPR 2025
它引用的顶会 Paper12
- PC-DARTS: Partial Channel Connections for Memory-Efficient Architecture SearchYuhui Xu, Lingxi Xie, Xiaopeng Zhang, Xin Chen 等ICLR 2020 · 被引用 691 次
- Understanding and Robustifying Differentiable Architecture SearchArber Zela, Thomas Elsken, Tonmoy Saikia, Yassine Marrakchi 等ICLR 2020 · 被引用 408 次
- Stabilizing Differentiable Architecture Search via Perturbation-based RegularizationXiangning Chen, Cho-Jui HsiehICML 2020 · 被引用 235 次
- Rethinking Architecture Selection in Differentiable NASRuochen Wang, Minhao Cheng, Xiangning Chen, Xiaocheng Tang 等ICLR 2021 · 被引用 213 次
- iDARTS: Differentiable Architecture Search with Stochastic Implicit GradientsMiao Zhang, Steven W. Su, Shirui Pan, Xiaojun Chang 等ICML 2021 · 被引用 81 次
相关 Paper
- β-DARTS: Beta-Decay Regularization for Differentiable Architecture SearchPeng Ye, Baopu Li, Yikang Li, Tao Chen 等CVPR 2022 · 被引用 106 次
- IS-DARTS: Stabilizing DARTS through Precise Measurement on Candidate ImportanceHongyi He, Longjun Liu, Haonan Zhang, Nanning ZhengAAAI 2024 · 被引用 21 次
- Interpreting Operation Selection in Differentiable Architecture Search: A Perspective from Influence-Directed ExplanationsMiao Zhang, Wei Huang, Bin YangNeurIPS 2022 · 被引用 7 次
- EC-DARTS: Inducing Equalized and Consistent Optimization into DARTSQinqin Zhou, Xiawu Zheng, Liujuan Cao, Bineng Zhong 等ICCV 2021 · 被引用 6 次
- Operation-Level Early Stopping for Robustifying Differentiable NASShen Jiang, Zipeng Ji, Guanghui Zhu, Chunfeng Yuan 等NeurIPS 2023 · 被引用 19 次
