Global Convergence of MAML and Theory-Inspired Neural Architecture Search for Few-Shot Learning
Haoxiang Wang, Yite Wang, Ruoyu Sun, Bo Li
Abstract
Model-agnostic meta-learning (MAML) and its variants have become popular approaches for few-shot learning. However, due to the non-convexity of deep neural nets (DNNs) and the bi-level formulation of MAML, the theoretical properties of MAML with DNNs remain largely unknown. In this paper, we first prove that MAML with over-parameterized DNNs is guaranteed to converge to global optima at a linear rate. Our convergence analysis indicates that MAML with over-parameterized DNNs is equivalent to kernel regression with a novel class of kernels, which we name as Meta Neural Tangent Kernels (MetaNTK). Then, we propose MetaNTK-NAS, a new training-free neural architecture search (NAS) method for few-shot learning that uses MetaNTK to rank and select architectures. Empirically, we compare our MetaNTK-NAS with previous NAS methods on two popular few-shot learning benchmarks, miniImageNet, and tieredImageNet. We show that the performance of MetaNTK-NAS is comparable or better than the state-of-the-art NAS method designed for few-shot learning while enjoying more than 100x speedup. We believe the efficiency of MetaNTK-NAS makes itself more practical for many real-world tasks. Our code is released at github.com/YiteWang/MetaNTK-NAS.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f1e59bda-0689-4ef1-87fe-99b04425fd8cCited by top-tier papers12
- Provable Domain Generalization via Invariant-Feature Subspace RecoveryHaoxiang Wang, Haozhe Si, Bo Li, Han ZhaoICML 2022 · 38 citations
- LEMON: Lossless model expansionYite Wang, Jiahao Su, Hanlin Lu, Cong Xie et al.ICLR 2024 · 25 citations
- Balanced Training for Sparse GANsYite Wang, Jing Wu, Naira Hovakimyan, Ruoyu SunNeurIPS 2023 · 16 citations
- Meta-Learning with Neural Bandit SchedulerYunzhe Qi, Yikun Ban, Tianxin Wei, Jiaru Zou et al.NeurIPS 2023 · 14 citations
- Meta-ticket: Finding optimal subnetworks for few-shot learning within randomly initialized neural networksDaiki Chijiwa, Shin'ya Yamaguchi, Atsutoshi Kumagai, Yasutoshi IdaNeurIPS 2022 · 12 citations
Builds on13
- Rapid Learning or Feature Reuse? Towards Understanding the Effectiveness of MAMLAniruddh Raghu, Maithra Raghu, Samy Bengio, Oriol VinyalsICLR 2020 · 736 citations
- Neural Architecture Search without TrainingJoe Mellor, Jack Turner, Amos Storkey, Elliot J. CrowleyICML 2021 · 477 citations
- Evaluating The Search Phase of Neural Architecture SearchKaicheng Yu, Christian Sciuto, Martin Jaggi, Claudiu Musat et al.ICLR 2020 · 370 citations
- Rethinking Architecture Selection in Differentiable NASRuochen Wang, Minhao Cheng, Xiangning Chen, Xiaocheng Tang et al.ICLR 2021 · 213 citations
- Bridging Multi-Task Learning and Meta-Learning: Towards Efficient Training and Effective AdaptationHaoxiang Wang, Han Zhao, Bo LiICML 2021 · 108 citations
Related papers
- Revisiting Neural Networks for Few-Shot Learning: A Zero-Cost NAS PerspectiveHaidong KangICML 2025
- Meta-Learning of Neural Architectures for Few-Shot LearningThomas Elsken, Benedikt Staffler, Jan Hendrik Metzen, Frank HutterCVPR 2020
- Rapid Model Architecture Adaption for Meta-LearningYiren Zhao, Xitong Gao, Ilia Shumailov, Nicolò Fusi et al.NeurIPS 2022 · 8 citations
- Towards Fast Adaptation of Neural Architectures with Meta LearningDongze Lian, Yin Zheng, Yintao Xu, Yanxiong Lu et al.ICLR 2020 · 95 citations
- Neural Architecture Search on ImageNet in Four GPU Hours: A Theoretically Inspired PerspectiveWuyang Chen, Xinyu Gong, Zhangyang WangICLR 2021 · 51 citations
