How to Train Your MAML to Excel in Few-Shot Classification
Han-Jia Ye, Wei-Lun Chao
摘要
Model-agnostic meta-learning (MAML) is arguably one of the most popular meta-learning algorithms nowadays. Nevertheless, its performance on few-shot classification is far behind many recent algorithms dedicated to the problem. In this paper, we point out several key facets of how to train MAML to excel in few-shot classification. First, we find that MAML needs a large number of gradient steps in its inner loop update, which contradicts its common usage in few-shot classification. Second, we find that MAML is sensitive to the class label assignments during meta-testing. Concretely, MAML meta-trains the initialization of an -way classifier. These ways, during meta-testing, then have""different permutations to be paired with a few-shot task of novel classes. We find that these permutations lead to a huge variance of accuracy, making MAML unstable in few-shot classification. Third, we investigate several approaches to make MAML permutation-invariant, among which meta-training a single vector to initialize all the weight vectors in the classification head performs the best. On benchmark datasets like MiniImageNet and TieredImageNet, our approach, which we name UNICORN-MAML, performs on a par with or even outperforms many recent few-shot classification algorithms, without sacrificing MAML's simplicity.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Bi-directional Feature Reconstruction Network for Fine-Grained Few-Shot Image ClassificationJijie Wu, Dongliang Chang, Aneeshan Sain, Xiaoxu Li 等AAAI 2023 · 被引用 76 次
- Channel Importance Matters in Few-Shot Image ClassificationXu Luo, Jing Xu, Zenglin XuICML 2022 · 被引用 57 次
- RankDNN: Learning to Rank for Few-Shot LearningQianyu Guo, Haotong Gong, Xujun Wei, Yanwei Fu 等AAAI 2023 · 被引用 27 次
- Scaling Few-Shot Learning for the Open WorldZhipeng Lin, Wenjing Yang, Haotian Wang, Haoang Chi 等AAAI 2024 · 被引用 5 次
- PERK: Long-Context Reasoning as Parameter-Efficient Test-Time LearningZeming Chen, Angelika Romanou, Gail Weiss, Antoine BosselutICLR 2026 · 被引用 4 次
它引用的顶会 Paper18
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Rapid Learning or Feature Reuse? Towards Understanding the Effectiveness of MAMLAniruddh Raghu, Maithra Raghu, Samy Bengio, Oriol VinyalsICLR 2020 · 被引用 736 次
- Meta-Dataset: A Dataset of Datasets for Learning to Learn from Few ExamplesEleni Triantafillou, Tyler Zhu, Vincent Dumoulin, Pascal Lamblin 等ICLR 2020 · 被引用 692 次
- A Baseline for Few-Shot Image ClassificationGuneet Singh Dhillon, Pratik Chaudhari, Avinash Ravichandran, Stefano SoattoICLR 2020 · 被引用 640 次
- Cross-Domain Few-Shot Classification via Learned Feature-Wise TransformationHung-Yu Tseng, Hsin-Ying Lee, Jia-Bin Huang, Ming-Hsuan YangICLR 2020 · 被引用 467 次
相关 Paper
- Repurposing Pretrained Models for Robust Out-of-domain Few-Shot LearningNamyeong Kwon, Hwidong Na, Gabriel Huang, Simon Lacoste-JulienICLR 2021 · 被引用 7 次
- OOD-MAML: Meta-Learning for Few-Shot Out-of-Distribution Detection and ClassificationTaewon Jeong, Heeyoung KimNeurIPS 2020 · 被引用 111 次
- A Nested Bi-level Optimization Framework for Robust Few Shot LearningKrishnaTeja Killamsetty, Changbin Li, Chen Zhao, Feng Chen 等AAAI 2022 · 被引用 12 次
- PLATINUM: Semi-Supervised Model Agnostic Meta-Learning using Submodular Mutual InformationChangbin Li, Suraj Kothawade, Feng Chen, Rishabh K. IyerICML 2022 · 被引用 6 次
- BOIL: Towards Representation Change for Few-shot LearningJaehoon Oh, Hyungjun Yoo, ChangHwan Kim, Se-Young YunICLR 2021 · 被引用 185 次
