Large-Scale Meta-Learning with Continual Trajectory Shifting
Jaewoong Shin, Haebeom Lee, Boqing Gong, Sung Ju Hwang
Abstract
Meta-learning of shared initialization parameters has shown to be highly effective in solving few-shot learning tasks. However, extending the framework to many-shot scenarios, which may further enhance its practicality, has been relatively overlooked due to the technical difficulties of meta-learning over long chains of inner-gradient steps. In this paper, we first show that allowing the meta-learners to take a larger number of inner gradient steps better captures the structure of heterogeneous and large-scale task distributions, thus results in obtaining better initialization points. Further, in order to increase the frequency of meta-updates even with the excessively long inner-optimization trajectories, we propose to estimate the required shift of the task-specific parameters with respect to the change of the initialization parameters. By doing so, we can arbitrarily increase the frequency of meta-updates and thus greatly improve the meta-level convergence as well as the quality of the learned initializations. We validate our method on a heterogeneous set of large-scale tasks and show that the algorithm largely outperforms the previous first-order meta-learning methods in terms of both generalization performance and convergence, as well as multi-task learning and fine-tuning baselines.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9273e591-89a9-4e26-80f6-72b5818845e2Cited by top-tier papers6
- How to Train Your MAML to Excel in Few-Shot ClassificationHan-Jia Ye, Wei-Lun ChaoICLR 2022 · 61 citations
- Memory Efficient Meta-Learning with Large ImagesJohn Bronskill, Daniela Massiceti, Massimiliano Patacchiola, Katja Hofmann et al.NeurIPS 2021 · 21 citations
- Sequential Reptile: Inter-Task Gradient Alignment for Multilingual LearningSeanie Lee, Haebeom Lee, Juho Lee, Sung Ju HwangICLR 2022 · 20 citations
- Meta-Learning with Self-Improving Momentum TargetJihoon Tack, Jongjin Park, Hankook Lee, Jaeho Lee et al.NeurIPS 2022 · 17 citations
- Learning Large-scale Neural Fields via Context Pruned Meta-LearningJihoon Tack, Subin Kim, Sihyun Yu, Jaeho Lee et al.NeurIPS 2023 · 16 citations
Builds on3
- A Baseline for Few-Shot Image ClassificationGuneet Singh Dhillon, Pratik Chaudhari, Avinash Ravichandran, Stefano SoattoICLR 2020 · 640 citations
- Meta-Learning with Warped Gradient DescentSebastian Flennerhag, Andrei A. Rusu, Razvan Pascanu, Francesco Visin et al.ICLR 2020 · 221 citations
- ES-MAML: Simple Hessian-Free Meta LearningXingyou Song, Wenbo Gao, Yuxiang Yang, Krzysztof Choromanski et al.ICLR 2020 · 128 citations
Related papers
- Meta-Learning with Adaptive HyperparametersSungyong Baik, Myungsub Choi, Janghoon Choi, Heewon Kim et al.NeurIPS 2020 · 164 citations
- Learning to Initialize: Can Meta Learning Improve Cross-task Generalization in Prompt Tuning?Chengwei Qin, Shafiq R. Joty, Qian Li, Ruochen ZhaoACL 2023 · 8 citations
- Meta-AdaM: An Meta-Learned Adaptive Optimizer with Momentum for Few-Shot LearningSiyuan Sun, Hongyang GaoNeurIPS 2023 · 51 citations
- On Enforcing Better Conditioned Meta-Learning for Rapid Few-Shot AdaptationMarkus Hiller, Mehrtash Harandi, Tom DrummondNeurIPS 2022 · 10 citations
- A Lazy Approach to Long-Horizon Gradient-Based Meta-LearningMuhammad Abdullah Jamal, Liqiang Wang, Boqing GongICCV 2021 · 9 citations
