Modular Meta-Learning with Shrinkage
Yutian Chen, Abram L. Friesen, Feryal M. P. Behbahani, Arnaud Doucet, David Budden, Matthew Hoffman, Nando de Freitas
摘要
Many real-world problems, including multi-speaker text-to-speech synthesis, can greatly benefit from the ability to meta-learn large models with only a few taskspecific components. Updating only these task-specific modules then allows the model to be adapted to low-data tasks for as many steps as necessary without risking overfitting. Unfortunately, existing meta-learning methods either do not scale to long adaptation or else rely on handcrafted task-specific architectures. Here, we propose a meta-learning approach that obviates the need for this often sub-optimal hand-selection. In particular, we develop general techniques based on Bayesian shrinkage to automatically discover and learn both task-specific and general reusable modules. Empirically, we demonstrate that our method discovers a small set of meaningful task-specific modules and outperforms existing metalearning approaches in domains like few-shot text-to-speech that have little task data and long adaptation horizons. We also show that existing meta-learning methods including MAML, iMAML, and Reptile emerge as special cases of our method. Automatically learning reusable and broadly applicable modular mechanisms is an open challenge in * Equal contribution. 34th Conference on Neural Information Processing Systems (NeurIPS 2020),
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- FedBABU: Toward Enhanced Representation for Federated Image ClassificationJaehoon Oh, Sangmook Kim, Se-Young YunICLR 2022 · 被引用 332 次
- BOIL: Towards Representation Change for Few-shot LearningJaehoon Oh, Hyungjun Yoo, ChangHwan Kim, Se-Young YunICLR 2021 · 被引用 185 次
- Learning where to learn: Gradient sparsity in meta and continual learningJohannes von Oswald, Dominic Zhao, Seijin Kobayashi, Simon Schug 等NeurIPS 2021 · 被引用 61 次
- Generalizing to New Physical Systems via Context-Informed Dynamics ModelMatthieu Kirchmeyer, Yuan Yin, Jérémie Donà, Nicolas Baskiotis 等ICML 2022 · 被引用 54 次
- A contrastive rule for meta-learningNicolas Zucchet, Simon Schug, Johannes von Oswald, Dominic Zhao 等NeurIPS 2022 · 被引用 22 次
它引用的顶会 Paper2
相关 Paper
- Towards Fast Adaptation of Neural Architectures with Meta LearningDongze Lian, Yin Zheng, Yintao Xu, Yanxiong Lu 等ICLR 2020 · 被引用 95 次
- Language-Agnostic Meta-Learning for Low-Resource Text-to-Speech with Articulatory FeaturesFlorian Lux, Ngoc Thang VuACL 2022 · 被引用 35 次
- M-NAS: Meta Neural Architecture SearchJiaxing Wang, Jiaxiang Wu, Haoli Bai, Jian ChengAAAI 2020 · 被引用 34 次
- Adversarial Meta Sampling for Multilingual Low-Resource Speech RecognitionYubei Xiao, Ke Gong, Pan Zhou, Guolin Zheng 等AAAI 2021 · 被引用 37 次
- Meta-Learning of Neural Architectures for Few-Shot LearningThomas Elsken, Benedikt Staffler, Jan Hendrik Metzen, Frank HutterCVPR 2020
