Revisiting Neural Networks for Few-Shot Learning: A Zero-Cost NAS Perspective
Haidong Kang
摘要
Neural Architecture Search (NAS) has recently outperformed hand-designed networks in various artificial intelligence areas. However, previous works only target a pre-defined task. For a new task in few-shot learning (FSL) scenarios, the architecture is either searched from scratch, which is neither efficient nor flexible, or borrowed architecture from the ones obtained on other tasks, which may lead to sub-optimal. Can we select the best neural architectures without involving any training and eliminate a significant portion of the search cost for new tasks in FSL? In this work, we provide an affirmative answer by proposing a novel information bottleneck (IB) theory driven Few-shot Neural Architecture Search (dubbed, IBFS) framework to address this issue. We first derive that the global convergence of Model-agnostic meta-learning (MAML) can be guaranteed by only considering the first-order loss landscape. Moreover, motivated by the observation that IB provides a unified view toward understanding machine learning models, we propose a novel Zero-Cost method tailored for FSL to rank and select architectures based on their expressivity obtained by IB mechanisms. Extensive experiments show that IBFS achieves state-of-the-art performance in FSL without training, which demonstrates the effectiveness of our IBFS.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- pTNAS: Progressive Neural Architecture Search for Tabular DataNaili Xing, Shaofeng Cai, Lingze Zeng, Jiaqi Zhu 等ICML 2026 · 被引用 4 次
- Revolutionizing Training-Free NAS: Towards Efficient Automatic Proxy Discovery via Large Language ModelsHaidong Kang, Lihong Lin, Hanling WangNeurIPS 2025 · 被引用 3 次
它引用的顶会 Paper40
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- NAS-Bench-201: Extending the Scope of Reproducible Neural Architecture SearchXuanyi Dong, Yi YangICLR 2020 · 被引用 825 次
- Rapid Learning or Feature Reuse? Towards Understanding the Effectiveness of MAMLAniruddh Raghu, Maithra Raghu, Samy Bengio, Oriol VinyalsICLR 2020 · 被引用 736 次
- Progressive Differentiable Architecture Search: Bridging the Depth Gap Between Search and EvaluationXin Chen, Lingxi Xie, Jun Wu, Qi TianICCV 2019 · 被引用 725 次
- PC-DARTS: Partial Channel Connections for Memory-Efficient Architecture SearchYuhui Xu, Lingxi Xie, Xiaopeng Zhang, Xin Chen 等ICLR 2020 · 被引用 691 次
相关 Paper
- Rapid Model Architecture Adaption for Meta-LearningYiren Zhao, Xitong Gao, Ilia Shumailov, Nicolò Fusi 等NeurIPS 2022 · 被引用 8 次
- Global Convergence of MAML and Theory-Inspired Neural Architecture Search for Few-Shot LearningHaoxiang Wang, Yite Wang, Ruoyu Sun, Bo LiCVPR 2022 · 被引用 37 次
- Towards Fast Adaptation of Neural Architectures with Meta LearningDongze Lian, Yin Zheng, Yintao Xu, Yanxiong Lu 等ICLR 2020 · 被引用 95 次
- M-NAS: Meta Neural Architecture SearchJiaxing Wang, Jiaxiang Wu, Haoli Bai, Jian ChengAAAI 2020 · 被引用 34 次
- Meta-Learning of Neural Architectures for Few-Shot LearningThomas Elsken, Benedikt Staffler, Jan Hendrik Metzen, Frank HutterCVPR 2020
