Meta-Learning of Neural Architectures for Few-Shot Learning
Thomas Elsken, Benedikt Staffler, Jan Hendrik Metzen, Frank Hutter
Abstract
The recent progress in neural architecture search (NAS) has allowed scaling the automated design of neural architectures to real-world domains, such as object detection and semantic segmentation. However, one prerequisite for the application of NAS are large amounts of labeled data and compute resources. This renders its application challenging in few-shot learning scenarios, where many related tasks need to be learned, each with limited amounts of data and compute time. Thus, few-shot learning is typically done with a fixed neural architecture. To improve upon this, we propose METANAS, the first method which fully integrates NAS with gradient-based meta-learning. METANAS optimizes a meta-architecture along with the meta-weights during meta-training. During meta-testing, architectures can be adapted to a novel task with a few steps of the task optimizer, that is: task adaptation becomes computationally cheap and requires only little data per task. Moreover, METANAS is agnostic in that it can be used with arbitrary model-agnostic meta-learning algorithms and arbitrary gradient-based NAS methods. Empirical results on standard few-shot classification benchmarks show that METANAS with a combination of DARTS and REPTILE yields state-of-the-art results.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5e48a357-890c-4fa5-87ff-9ef3ecc870c1Cited by top-tier papers26
- Partial Is Better Than All: Revisiting Fine-tuning Strategy for Few-shot LearningZhiqiang Shen, Zechun Liu, Jie Qin, Marios Savvides et al.AAAI 2021 · 203 citations
- Federated Hyperparameter Tuning: Challenges, Baselines, and Connections to Weight-SharingMikhail Khodak, Renbo Tu, Tian Li, Liam Li et al.NeurIPS 2021 · 111 citations
- Parameter Prediction for Unseen Deep ArchitecturesBoris Knyazev, Michal Drozdzal, Graham W. Taylor, Adriana Romero-SorianoNeurIPS 2021 · 111 citations
- Rapid Neural Architecture Search by Learning to Generate Graphs from DatasetsHayeon Lee, Eunyoung Hyung, Sung Ju HwangICLR 2021 · 57 citations
- Meta Navigator: Search for a Good Adaptation Policy for Few-shot LearningChi Zhang, Henghui Ding, Guosheng Lin, Ruibo Li et al.ICCV 2021 · 51 citations
Builds on3
- Understanding and Robustifying Differentiable Architecture SearchArber Zela, Thomas Elsken, Tonmoy Saikia, Yassine Marrakchi et al.ICLR 2020 · 408 citations
- Towards Fast Adaptation of Neural Architectures with Meta LearningDongze Lian, Yin Zheng, Yintao Xu, Yanxiong Lu et al.ICLR 2020 · 95 citations
- AutoDispNet: Improving Disparity Estimation With AutoMLTonmoy Saikia, Yassine Marrakchi, Arber Zela, Frank Hutter et al.ICCV 2019 · 84 citations
Related papers
- Rapid Model Architecture Adaption for Meta-LearningYiren Zhao, Xitong Gao, Ilia Shumailov, Nicolò Fusi et al.NeurIPS 2022 · 8 citations
- M-NAS: Meta Neural Architecture SearchJiaxing Wang, Jiaxiang Wu, Haoli Bai, Jian ChengAAAI 2020 · 34 citations
- Global Convergence of MAML and Theory-Inspired Neural Architecture Search for Few-Shot LearningHaoxiang Wang, Yite Wang, Ruoyu Sun, Bo LiCVPR 2022 · 37 citations
- Neural Fine-Tuning Search for Few-Shot LearningPanagiotis Eustratiadis, Lukasz Dudziak, Da Li, Timothy M. HospedalesICLR 2024 · 7 citations
- Revisiting Neural Networks for Few-Shot Learning: A Zero-Cost NAS PerspectiveHaidong KangICML 2025
