How to Distribute Data across Tasks for Meta-Learning?
Alexandru Cioba, Michael Bromberg, Qian Wang, Ritwik Niyogi, Georgios Batzolis, Jezabel R. Garcia, Da-Shan Shiu, Alberto Bernacchia
Abstract
Meta-learning models transfer the knowledge acquired from previous tasks to quickly learn new ones. They are trained on benchmarks with a fixed number of data points per task. This number is usually arbitrary and it is unknown how it affects performance at testing. Since labelling of data is expensive, finding the optimal allocation of labels across training tasks may reduce costs. Given a fixed budget of labels, should we use a small number of highly labelled tasks, or many tasks with few labels each? Should we allocate more labels to some tasks and less to others? We show that: 1) If tasks are homogeneous, there is a uniform optimal allocation, whereby all tasks get the same amount of data; 2) At fixed budget, there is a trade-off between number of tasks and number of data points per task, with a unique solution for the optimum; 3) When trained separately, harder task should get more data, at the cost of a smaller number of tasks; 4) When training on a mixture of easy and hard tasks, more data should be allocated to easy tasks. Interestingly, Neuroscience experiments have shown that human visual skills also transfer better from easy tasks. We prove these results mathematically on mixed linear regression, and we show empirically that the same results hold for few-shot image classification on CIFAR-FS and mini-ImageNet. Our results provide guidance for allocating labels across tasks when collecting data for meta-learning.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a3e66cfa-250c-4a12-8c9c-50e8fc2f9d06Cited by top-tier papers2
- Dual-Level Curriculum Meta-Learning for Noisy Few-Shot Learning TasksXiaofan Que, Qi YuAAAI 2024 · 5 citations
- Is Meta-Learning Out? Rethinking Unsupervised Few-Shot Classification with Limited EntropyYunchuan Guan, Yu Liu, Ke Zhou, Zhiqi Shen et al.ICCV 2025 · 2 citations
Builds on4
- Meta-learning for Mixed Linear RegressionWeihao Kong, Raghav Somani, Zhao Song, Sham M. Kakade et al.ICML 2020 · 70 citations
- How Important is the Train-Validation Split in Meta-Learning?Yu Bai, Minshuo Chen, Pan Zhou, Tuo Zhao et al.ICML 2021 · 60 citations
- Modeling and Optimization Trade-off in Meta-learningKatelyn Gao, Ozan SenerNeurIPS 2020 · 33 citations
- Meta-learning with negative learning ratesAlberto BernacchiaICLR 2021 · 4 citations
Related papers
- Meta-Learning without MemorizationMingzhang Yin, George Tucker, Mingyuan Zhou, Sergey Levine et al.ICLR 2020 · 201 citations
- Few-Round Learning for Federated LearningYounghyun Park, Dong-Jun Han, Do-Yeon Kim, Jun Seo et al.NeurIPS 2021 · 31 citations
- Learning to Balance: Bayesian Meta-Learning for Imbalanced and Out-of-distribution TasksHaebeom Lee, Hayeon Lee, Donghyun Na, Saehoon Kim et al.ICLR 2020 · 115 citations
- Meta Omnium: A Benchmark for General-Purpose Learning-to-LearnOndrej Bohdal, Yinbing Tian, Yongshuo Zong, Ruchika Chavhan et al.CVPR 2023
- A Theoretical Analysis of the Number of Shots in Few-Shot LearningTianshi Cao, Marc T. Law, Sanja FidlerICLR 2020 · 75 citations
