BOIL: Towards Representation Change for Few-shot Learning
Jaehoon Oh, Hyungjun Yoo, ChangHwan Kim, Se-Young Yun
摘要
Model Agnostic Meta-Learning (MAML) is one of the most representative of gradient-based meta-learning algorithms. MAML learns new tasks with a few data samples using inner updates from a meta-initialization point and learns the meta-initialization parameters with outer updates. It has recently been hypothesized that representation reuse, which makes little change in efficient representations, is the dominant factor in the performance of the meta-initialized model through MAML in contrast to representation change, which causes a significant change in representations. In this study, we investigate the necessity of representation change for the ultimate goal of few-shot learning, which is solving domain-agnostic tasks. To this aim, we propose a novel meta-learning algorithm, called BOIL (Body Only update in Inner Loop), which updates only the body (extractor) of the model and freezes the head (classifier) during inner loop updates. BOIL leverages representation change rather than representation reuse. This is because feature vectors (representations) have to move quickly to their corresponding frozen head vectors. We visualize this property using cosine similarity, CKA, and empirical results without the head. BOIL empirically shows significant performance improvement over MAML, particularly on cross-domain tasks. The results imply that representation change in gradient-based meta-learning approaches is a critical component.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper33
- FedBABU: Toward Enhanced Representation for Federated Image ClassificationJaehoon Oh, Sangmook Kim, Se-Young YunICLR 2022 · 被引用 332 次
- Meta-Learning with Task-Adaptive Loss Function for Few-Shot LearningSungyong Baik, Janghoon Choi, Heewon Kim, Dohee Cho 等ICCV 2021 · 被引用 146 次
- Matching Feature Sets for Few-Shot Image ClassificationArman Afrasiyabi, Hugo Larochelle, Jean-François Lalonde, Christian GagnéCVPR 2022 · 被引用 124 次
- Meta-DMoE: Adapting to Domain Shift by Meta-Distillation from Mixture-of-ExpertsTao Zhong, Zhixiang Chi, Li Gu, Yang Wang 等NeurIPS 2022 · 被引用 70 次
- Understanding Cross-Domain Few-Shot Learning Based on Domain Similarity and Few-Shot DifficultyJaehoon Oh, Sungnyun Kim, Namgyu Ho, Jin-Hwa Kim 等NeurIPS 2022 · 被引用 69 次
它引用的顶会 Paper8
- Rapid Learning or Feature Reuse? Towards Understanding the Effectiveness of MAMLAniruddh Raghu, Maithra Raghu, Samy Bengio, Oriol VinyalsICLR 2020 · 被引用 736 次
- Meta-Dataset: A Dataset of Datasets for Learning to Learn from Few ExamplesEleni Triantafillou, Tyler Zhu, Vincent Dumoulin, Pascal Lamblin 等ICLR 2020 · 被引用 692 次
- Cross-Domain Few-Shot Classification via Learned Feature-Wise TransformationHung-Yu Tseng, Hsin-Ying Lee, Jia-Bin Huang, Ming-Hsuan YangICLR 2020 · 被引用 467 次
- Meta-Learning with Warped Gradient DescentSebastian Flennerhag, Andrei A. Rusu, Razvan Pascanu, Francesco Visin 等ICLR 2020 · 被引用 221 次
- Meta-Learning without MemorizationMingzhang Yin, George Tucker, Mingyuan Zhou, Sergey Levine 等ICLR 2020 · 被引用 201 次
相关 Paper
- How to Train Your MAML to Excel in Few-Shot ClassificationHan-Jia Ye, Wei-Lun ChaoICLR 2022 · 被引用 61 次
- Meta-Learning with Adaptive HyperparametersSungyong Baik, Myungsub Choi, Janghoon Choi, Heewon Kim 等NeurIPS 2020 · 被引用 164 次
- Meta-Learning with a Geometry-Adaptive PreconditionerSuhyun Kang, Duhun Hwang, Moonjung Eo, Taesup Kim 等CVPR 2023
- How Fine-Tuning Allows for Effective Meta-LearningKurtland Chua, Qi Lei, Jason D. LeeNeurIPS 2021 · 被引用 57 次
- A Nested Bi-level Optimization Framework for Robust Few Shot LearningKrishnaTeja Killamsetty, Changbin Li, Chen Zhao, Feng Chen 等AAAI 2022 · 被引用 12 次
