Two Sides of Meta-Learning Evaluation: In vs. Out of Distribution
Amrith Setlur, Oscar Li, Virginia Smith
Abstract
We categorize meta-learning evaluation into two settings: [ID], in which the train and test tasks are sampled from the same underlying task distribution, and [OOD], in which they are not. While most meta-learning theory and some FSL applications follow the ID setting, we identify that most existing few-shot classification benchmarks instead reflect OOD evaluation, as they use disjoint sets of train (base) and test (novel) classes for task generation. This discrepancy is problematic because -- as we show on numerous benchmarks -- meta-learning methods that perform better on existing OOD datasets may perform significantly worse in the ID setting. In addition, in the OOD setting, even though current FSL benchmarks seem befitting, our study highlights concerns in 1) reliably performing model selection for a given meta-learning method, and 2) consistently comparing the performance of different methods. To address these concerns, we provide suggestions on how to construct FSL benchmarks to allow for ID evaluation as well as more reliable OOD evaluation. Our work aims to inform the meta-learning community about the importance and distinction of ID vs. OOD evaluation, as well as the subtleties of OOD evaluation with current benchmarks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c3e25dc1-e048-466a-b7bd-1b38e36e12b5Cited by top-tier papers1
Ask how each one uses itBuilds on7
- Personalized Federated Learning with Theoretical Guarantees: A Model-Agnostic Meta-Learning ApproachAlireza Fallah, Aryan Mokhtari, Asuman E. OzdaglarNeurIPS 2020 · 1,354 citations
- Meta-Dataset: A Dataset of Datasets for Learning to Learn from Few ExamplesEleni Triantafillou, Tyler Zhu, Vincent Dumoulin, Pascal Lamblin et al.ICLR 2020 · 692 citations
- Learning to Balance: Bayesian Meta-Learning for Imbalanced and Out-of-distribution TasksHaebeom Lee, Hayeon Lee, Donghyun Na, Saehoon Kim et al.ICLR 2020 · 115 citations
- Learning a Universal Template for Few-shot Dataset GeneralizationEleni Triantafillou, Hugo Larochelle, Richard S. Zemel, Vincent DumoulinICML 2021 · 113 citations
- Generalization of Model-Agnostic Meta-Learning Algorithms: Recurring and Unseen TasksAlireza Fallah, Aryan Mokhtari, Asuman E. OzdaglarNeurIPS 2021 · 63 citations
Related papers
- Characterizing Generalization under Out-Of-Distribution Shifts in Deep Metric LearningTimo Milbich, Karsten Roth, Samarth Sinha, Ludwig Schmidt et al.NeurIPS 2021 · 26 citations
- OOD-MAML: Meta-Learning for Few-Shot Out-of-Distribution Detection and ClassificationTaewon Jeong, Heeyoung KimNeurIPS 2020 · 111 citations
- MetaCoCo: A New Few-Shot Classification Benchmark with Spurious CorrelationMin Zhang, Haoxuan Li, Fei Wu, Kun KuangICLR 2024 · 18 citations
- Meta OOD Learning For Continuously Adaptive OOD DetectionXinheng Wu, Jie Lu, Zhen Fang, Guangquan ZhangICCV 2023 · 15 citations
- MetaOOD: Automatic Selection of OOD Detection ModelsYuehan Qin, Yichi Zhang, Yi Nian, Xueying Ding et al.ICLR 2025
