iDARTS: Differentiable Architecture Search with Stochastic Implicit Gradients
Miao Zhang, Steven W. Su, Shirui Pan, Xiaojun Chang, M. Ehsan Abbasnejad, Reza Haffari
摘要
Differentiable ARchiTecture Search (DARTS) has recently become the mainstream of neural architecture search (NAS) due to its efficiency and simplicity. With a gradient-based bi-level optimization, DARTS alternately optimizes the inner model weights and the outer architecture parameter in a weight-sharing supernet. A key challenge to the scalability and quality of the learned architectures is the need for differentiating through the inner-loop optimisation. While much has been discussed about several potentially fatal factors in DARTS, the architecture gradient, a.k.a. hypergradient, has received less attention. In this paper, we tackle the hypergradient computation in DARTS based on the implicit function theorem, making it only depends on the obtained solution to the inner-loop optimization and agnostic to the optimization path. To further reduce the computational requirements, we formulate a stochastic hypergradient approximation for differentiable NAS, and theoretically show that the architecture optimization with the proposed method, named iDARTS, is expected to converge to a stationary point. Comprehensive experiments on two NAS benchmark search spaces and the common NAS search space verify the effectiveness of our proposed method. It leads to architectures outperforming, with large margins, those learned by the baseline methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Making Scalable Meta Learning PracticalSang Keun Choe, Sanket Vaibhav Mehta, Hwijeen Ahn, Willie Neiswanger 等NeurIPS 2023 · 被引用 28 次
- Multiset-Equivariant Set Prediction with Approximate Implicit DifferentiationYan Zhang, David W. Zhang, Simon Lacoste-Julien, Gertjan J. Burghouts 等ICLR 2022 · 被引用 22 次
- Deep invariant networks with differentiable augmentation layersCédric Rommel, Thomas Moreau, Alexandre GramfortNeurIPS 2022 · 被引用 11 次
- Weighted Mutual Learning with Diversity-Driven Model CompressionMiao Zhang, Li Wang, David Campos, Wei Huang 等NeurIPS 2022 · 被引用 10 次
- Meta-learning Adaptive Deep Kernel Gaussian Processes for Molecular Property PredictionWenlin Chen, Austin Tripp, José Miguel Hernández-LobatoICLR 2023 · 被引用 8 次
它引用的顶会 Paper17
- NAS-Bench-201: Extending the Scope of Reproducible Neural Architecture SearchXuanyi Dong, Yi YangICLR 2020 · 被引用 825 次
- Progressive Differentiable Architecture Search: Bridging the Depth Gap Between Search and EvaluationXin Chen, Lingxi Xie, Jun Wu, Qi TianICCV 2019 · 被引用 725 次
- PC-DARTS: Partial Channel Connections for Memory-Efficient Architecture SearchYuhui Xu, Lingxi Xie, Xiaopeng Zhang, Xin Chen 等ICLR 2020 · 被引用 691 次
- Hierarchical Neural Architecture Search for Deep Stereo MatchingXuelian Cheng, Yiran Zhong, Mehrtash Harandi, Yuchao Dai 等NeurIPS 2020 · 被引用 436 次
- Understanding and Robustifying Differentiable Architecture SearchArber Zela, Thomas Elsken, Tonmoy Saikia, Yassine Marrakchi 等ICLR 2020 · 被引用 408 次
相关 Paper
- -DARTS: Mitigating Performance Collapse by Harmonizing Operation Selection among CellsSajad Movahedi, Melika Adabinejad, Ayyoob Imani, Arezou Keshavarz 等ICLR 2023
- Interpreting Operation Selection in Differentiable Architecture Search: A Perspective from Influence-Directed ExplanationsMiao Zhang, Wei Huang, Bin YangNeurIPS 2022 · 被引用 7 次
- IS-DARTS: Stabilizing DARTS through Precise Measurement on Candidate ImportanceHongyi He, Longjun Liu, Haonan Zhang, Nanning ZhengAAAI 2024 · 被引用 21 次
- Shapley-NAS: Discovering Operation Contribution for Neural Architecture SearchHan Xiao, Ziwei Wang, Zheng Zhu, Jie Zhou 等CVPR 2022 · 被引用 59 次
- Rethinking Architecture Selection in Differentiable NASRuochen Wang, Minhao Cheng, Xiangning Chen, Xiaocheng Tang 等ICLR 2021 · 被引用 213 次
