Modular Meta-Learning with Shrinkage
Yutian Chen, Abram L. Friesen, Feryal M. P. Behbahani, Arnaud Doucet, David Budden, Matthew Hoffman, Nando de Freitas
Abstract
Many real-world problems, including multi-speaker text-to-speech synthesis, can greatly benefit from the ability to meta-learn large models with only a few taskspecific components. Updating only these task-specific modules then allows the model to be adapted to low-data tasks for as many steps as necessary without risking overfitting. Unfortunately, existing meta-learning methods either do not scale to long adaptation or else rely on handcrafted task-specific architectures. Here, we propose a meta-learning approach that obviates the need for this often sub-optimal hand-selection. In particular, we develop general techniques based on Bayesian shrinkage to automatically discover and learn both task-specific and general reusable modules. Empirically, we demonstrate that our method discovers a small set of meaningful task-specific modules and outperforms existing metalearning approaches in domains like few-shot text-to-speech that have little task data and long adaptation horizons. We also show that existing meta-learning methods including MAML, iMAML, and Reptile emerge as special cases of our method. Automatically learning reusable and broadly applicable modular mechanisms is an open challenge in * Equal contribution. 34th Conference on Neural Information Processing Systems (NeurIPS 2020),
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext cf920852-cae4-4a3c-9bbd-64dd6f2778cbCited by top-tier papers10
- FedBABU: Toward Enhanced Representation for Federated Image ClassificationJaehoon Oh, Sangmook Kim, Se-Young YunICLR 2022 · 332 citations
- BOIL: Towards Representation Change for Few-shot LearningJaehoon Oh, Hyungjun Yoo, ChangHwan Kim, Se-Young YunICLR 2021 · 185 citations
- Learning where to learn: Gradient sparsity in meta and continual learningJohannes von Oswald, Dominic Zhao, Seijin Kobayashi, Simon Schug et al.NeurIPS 2021 · 61 citations
- Generalizing to New Physical Systems via Context-Informed Dynamics ModelMatthieu Kirchmeyer, Yuan Yin, Jérémie Donà, Nicolas Baskiotis et al.ICML 2022 · 54 citations
- A contrastive rule for meta-learningNicolas Zucchet, Simon Schug, Johannes von Oswald, Dominic Zhao et al.NeurIPS 2022 · 22 citations
Builds on2
- Meta-Learning with Warped Gradient DescentSebastian Flennerhag, Andrei A. Rusu, Razvan Pascanu, Francesco Visin et al.ICLR 2020 · 221 citations
- Learning to Balance: Bayesian Meta-Learning for Imbalanced and Out-of-distribution TasksHaebeom Lee, Hayeon Lee, Donghyun Na, Saehoon Kim et al.ICLR 2020 · 115 citations
Related papers
- Towards Fast Adaptation of Neural Architectures with Meta LearningDongze Lian, Yin Zheng, Yintao Xu, Yanxiong Lu et al.ICLR 2020 · 95 citations
- Language-Agnostic Meta-Learning for Low-Resource Text-to-Speech with Articulatory FeaturesFlorian Lux, Ngoc Thang VuACL 2022 · 35 citations
- M-NAS: Meta Neural Architecture SearchJiaxing Wang, Jiaxiang Wu, Haoli Bai, Jian ChengAAAI 2020 · 34 citations
- Adversarial Meta Sampling for Multilingual Low-Resource Speech RecognitionYubei Xiao, Ke Gong, Pan Zhou, Guolin Zheng et al.AAAI 2021 · 37 citations
- Meta-Learning of Neural Architectures for Few-Shot LearningThomas Elsken, Benedikt Staffler, Jan Hendrik Metzen, Frank HutterCVPR 2020
