Provable Generalization of Overparameterized Meta-learning Trained with SGD
Yu Huang, Yingbin Liang, Longbo Huang
Abstract
Despite the superior empirical success of deep meta-learning, theoretical understanding of overparameterized meta-learning is still limited. This paper studies the generalization of a widely used meta-learning approach, Model-Agnostic Meta-Learning (MAML), which aims to find a good initialization for fast adaptation to new tasks. Under a mixed linear regression model, we analyze the generalization properties of MAML trained with SGD in the overparameterized regime. We provide both upper and lower bounds for the excess risk of MAML, which captures how SGD dynamics affect these generalization bounds. With such sharp characterizations, we further explore how various learning parameters impact the generalization capability of overparameterized MAML, including explicitly identifying typical data and task distributions that can achieve diminishing generalization error with overparameterization, and characterizing the impact of adaptation learning rate on both excess risk and the early stopping time. Our theoretical findings are further validated by experiments.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2dfea777-d5a2-4f9a-9b1f-408bae6a65c6Cited by top-tier papers5
- Understanding Benign Overfitting in Gradient-Based Meta LearningLisha Chen, Songtao Lu, Tianyi ChenNeurIPS 2022 · 20 citations
- Online Constrained Meta-Learning: Provable Guarantees for GeneralizationSiyuan Xu, Minghui ZhuNeurIPS 2023 · 10 citations
- Meta-Reinforcement Learning with Universal Policy Adaptation: Provable Near-Optimality under All-task Optimum ComparatorSiyuan Xu, Minghui ZhuNeurIPS 2024 · 8 citations
- Theoretical Characterization of the Generalization Performance of Overfitted Meta-LearningPeizhong Ju, Yingbin Liang, Ness B. ShroffICLR 2023 · 3 citations
- FedMeNF: Privacy-Preserving Federated Meta-Learning for Neural FieldsJunhyeog Yun, Minui Hong, Gunhee KimICCV 2025
Builds on13
- Few-shot Text Classification with Distributional SignaturesYujia Bao, Menghua Wu, Shiyu Chang, Regina BarzilayICLR 2020 · 183 citations
- Meta-learning for Mixed Linear RegressionWeihao Kong, Raghav Somani, Zhao Song, Sham M. Kakade et al.ICML 2020 · 70 citations
- Generalization Bounds For Meta-Learning: An Information-Theoretic AnalysisQi Chen, Changjian Shui, Mario MarchandNeurIPS 2021 · 66 citations
- Generalization of Model-Agnostic Meta-Learning Algorithms: Recurring and Unseen TasksAlireza Fallah, Aryan Mokhtari, Asuman E. OzdaglarNeurIPS 2021 · 63 citations
- How Important is the Train-Validation Split in Meta-Learning?Yu Bai, Minshuo Chen, Pan Zhou, Tuo Zhao et al.ICML 2021 · 60 citations
Related papers
- Unraveling Model-Agnostic Meta-Learning via The Adaptation Learning RateYingtian Zou, Fusheng Liu, Qianxiao LiICLR 2022 · 11 citations
- Meta-learning with negative learning ratesAlberto BernacchiaICLR 2021 · 4 citations
- Task-Robust Model-Agnostic Meta-LearningLiam Collins, Aryan Mokhtari, Sanjay ShakkottaiNeurIPS 2020 · 66 citations
- Risk Bounds of Accelerated SGD for Overparameterized Linear RegressionXuheng Li, Yihe Deng, Jingfeng Wu, Dongruo Zhou et al.ICLR 2024 · 7 citations
- Towards Sample-efficient Overparameterized Meta-learningYue Sun, Adhyyan Narang, Halil Ibrahim Gulluk, Samet Oymak et al.NeurIPS 2021 · 26 citations
