Learning to Forget for Meta-Learning
Sungyong Baik, Seokil Hong, Kyoung Mu Lee
Abstract
Few-shot learning is a challenging problem where the goal is to achieve generalization from only few examples. Model-agnostic meta-learning (MAML) tackles the problem by formulating prior knowledge as a common initialization across tasks, which is then used to quickly adapt to unseen tasks. However, forcibly sharing an initialization can lead to conflicts among tasks and the compromised (undesired by tasks) location on optimization landscape, thereby hindering the task adaptation. Further, we observe that the degree of conflict differs among not only tasks but also layers of a neural network. Thus, we propose task-and-layer-wise attenuation on the compromised initialization to reduce its influence. As the attenuation dynamically controls (or selectively forgets) the influence of prior knowledge for a given task and each layer, we name our method as L2F (Learn to Forget) 1 . The experimental results demonstrate that the proposed method provides faster adaptation and greatly improves the performance. Furthermore, L2F can be easily applied and improve other state-of-the-art MAML-based frameworks, illustrating its simplicity and generalizability.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 78d541b2-8b4b-41d9-8e8d-40e6d91fed46Cited by top-tier papers16
- Meta-Learning with Adaptive HyperparametersSungyong Baik, Myungsub Choi, Janghoon Choi, Heewon Kim et al.NeurIPS 2020 · 164 citations
- Meta-Learning with Task-Adaptive Loss Function for Few-Shot LearningSungyong Baik, Janghoon Choi, Heewon Kim, Dohee Cho et al.ICCV 2021 · 146 citations
- Curvature Generation in Curved Spaces for Few-Shot LearningZhi Gao, Yuwei Wu, Yunde Jia, Mehrtash HarandiICCV 2021 · 71 citations
- Sketch3T: Test-Time Training for Zero-Shot SBIRAneeshan Sain, Ayan Kumar Bhunia, Vaishnav Potlapalli, Pinaki Nath Chowdhury et al.CVPR 2022 · 55 citations
- Revisit Multimodal Meta-Learning through the Lens of Multi-Task LearningMilad Abdollahzadeh, Touba Malekzadeh, Ngai-Man CheungNeurIPS 2021 · 39 citations
Builds on1
Related papers
- Rapid Model Architecture Adaption for Meta-LearningYiren Zhao, Xitong Gao, Ilia Shumailov, Nicolò Fusi et al.NeurIPS 2022 · 8 citations
- On Fast Adversarial Robustness Adaptation in Model-Agnostic Meta-LearningRen Wang, Kaidi Xu, Sijia Liu, Pin-Yu Chen et al.ICLR 2021 · 17 citations
- MATE: Plugging in Model Awareness to Task Embedding for Meta LearningXiaohan Chen, Zhangyang Wang, Siyu Tang, Krikamol MuandetNeurIPS 2020 · 10 citations
- Enabling Few-Shot Learning with PID Control: A Layer Adaptive OptimizerLe Yu, Xinde Li, Pengfei Zhang, Zhentong Zhang et al.ICML 2024 · 3 citations
- OOD-MAML: Meta-Learning for Few-Shot Out-of-Distribution Detection and ClassificationTaewon Jeong, Heeyoung KimNeurIPS 2020 · 111 citations
