Meta-Learning via Learning with Distributed Memory
Sudarshan Babu, Pedro Savarese, Michael Maire
Abstract
We demonstrate that efficient meta-learning can be achieved via end-to-end training of deep neural networks with memory distributed across layers. The persistent state of this memory assumes the entire burden of guiding task adaptation. Moreover, its distributed nature is instrumental in orchestrating adaptation. Ablation experiments demonstrate that providing relevant feedback to memory units distributed across the depth of the network enables them to guide adaptation throughout the entire network. Our results show that this is a successful strategy for simplifying metalearning -often cast as a bi-level optimization problem -to standard end-to-end training, while outperforming gradient-based, prototype-based, and other memorybased meta-learning strategies. Additionally, our adaptation strategy naturally handles online learning scenarios with a significant delay between observing a sample and its corresponding label -a setting in which other approaches struggle. Adaptation via distributed memory is effective across a wide range of learning tasks, ranging from classification to online few-shot semantic segmentation.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on5
- Meta-Learning with Warped Gradient DescentSebastian Flennerhag, Andrei A. Rusu, Razvan Pascanu, Francesco Visin et al.ICLR 2020 · 221 citations
- Meta Learning Backpropagation And Improving ItLouis Kirsch, Jürgen SchmidhuberNeurIPS 2021 · 70 citations
- Meta-Learning Deep Energy-Based Memory ModelsSergey Bartunov, Jack W. Rae, Simon Osindero, Timothy P. LillicrapICLR 2020 · 35 citations
- Wandering within a world: Online contextualized few-shot learningMengye Ren, Michael Louis Iuzzolino, Michael Curtis Mozer, Richard S. ZemelICLR 2021 · 33 citations
- FSS-1000: A 1000-Class Dataset for Few-Shot SegmentationXiang Li, Tianhan Wei, Yau Pun Chen, Yu-Wing Tai et al.CVPR 2020
Related papers
- Hierarchical Variational Memory for Few-shot Learning Across DomainsYing-Jun Du, Xiantong Zhen, Ling Shao, Cees G. M. SnoekICLR 2022 · 24 citations
- p-Meta: Towards On-device Deep Model AdaptationZhongnan Qu, Zimu Zhou, Yongxin Tong, Lothar ThieleKDD 2022 · 9 citations
- MetaNorm: Learning to Normalize Few-Shot Batches Across DomainsYing-Jun Du, Xiantong Zhen, Ling Shao, Cees G. M. SnoekICLR 2021 · 26 citations
- Learning to Learn Variational Semantic MemoryXiantong Zhen, Ying-Jun Du, Huan Xiong, Qiang Qiu et al.NeurIPS 2020 · 40 citations
- Meta-Learning of Neural Architectures for Few-Shot LearningThomas Elsken, Benedikt Staffler, Jan Hendrik Metzen, Frank HutterCVPR 2020
