A contrastive rule for meta-learning
Nicolas Zucchet, Simon Schug, Johannes von Oswald, Dominic Zhao, João Sacramento
摘要
Humans and other animals are capable of improving their learning performance as they solve related tasks from a given problem domain, to the point of being able to learn from extremely limited data. While synaptic plasticity is generically thought to underlie learning in the brain, the precise neural and synaptic mechanisms by which learning processes improve through experience are not well understood. Here, we present a general-purpose, biologically-plausible meta-learning rule which estimates gradients with respect to the parameters of an underlying learning algorithm by simply running it twice. Our rule may be understood as a generalization of contrastive Hebbian learning to meta-learning and notably, it neither requires computing second derivatives nor going backwards in time, two characteristic features of previous gradient-based methods that are hard to conceive in physical neural circuits. We demonstrate the generality of our rule by applying it to two distinct models: a complex synapse with internal states which consolidate task-shared information, and a dual-system architecture in which a primary network is rapidly modulated by another one to learn the specifics of each task. For both models, our meta-learning rule matches or outperforms reference algorithms on a wide range of benchmark problems, while only using information presumed to be locally available at neurons and synapses. We corroborate these findings with a theoretical analysis of the gradient estimation error incurred by our rule. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Holomorphic Equilibrium Propagation Computes Exact Gradients Through Finite Size OscillationsAxel Laborieux, Friedemann ZenkeNeurIPS 2022 · 被引用 65 次
- Dataset Distillation with Convexified Implicit GradientsNoel Loo, Ramin M. Hasani, Mathias Lechner, Daniela RusICML 2023 · 被引用 56 次
- Energy-based learning algorithms for analog computing: a comparative studyBenjamin Scellier, Maxence Ernoult, Jack D. Kendall, Suhas KumarNeurIPS 2023 · 被引用 54 次
- The least-control principle for local learning at equilibriumAlexander Meulemans, Nicolas Zucchet, Seijin Kobayashi, Johannes von Oswald 等NeurIPS 2022 · 被引用 32 次
- Making Scalable Meta Learning PracticalSang Keun Choe, Sanket Vaibhav Mehta, Hwijeen Ahn, Willie Neiswanger 等NeurIPS 2023 · 被引用 28 次
它引用的顶会 Paper5
- BatchEnsemble: an Alternative Approach to Efficient Ensemble and Lifelong LearningYeming Wen, Dustin Tran, Jimmy BaICLR 2020 · 被引用 569 次
- Continual learning with hypernetworksJohannes von Oswald, Christian Henning, João Sacramento, Benjamin F. GreweICLR 2020 · 被引用 412 次
- Look-ahead Meta Learning for Continual LearningGunshi Gupta, Karmesh Yadav, Liam PaullNeurIPS 2020 · 被引用 74 次
- Learning where to learn: Gradient sparsity in meta and continual learningJohannes von Oswald, Dominic Zhao, Seijin Kobayashi, Simon Schug 等NeurIPS 2021 · 被引用 61 次
- Modular Meta-Learning with ShrinkageYutian Chen, Abram L. Friesen, Feryal M. P. Behbahani, Arnaud Doucet 等NeurIPS 2020 · 被引用 35 次
相关 Paper
- Learning to Learn with Feedback and Local PlasticityJack Lindsey, Ashok Litwin-KumarNeurIPS 2020 · 被引用 38 次
- A meta-learning approach to (re)discover plasticity rules that carve a desired function into a neural networkBasile Confavreux, Friedemann Zenke, Everton J. Agnes, Timothy P. Lillicrap 等NeurIPS 2020 · 被引用 40 次
- Meta-Reinforcement Learning with Self-Modifying NetworksMathieu Chalvidal, Thomas Serre, Rufin VanRullenNeurIPS 2022 · 被引用 13 次
- Meta-Learning Bidirectional Update RulesMark Sandler, Max Vladymyrov, Andrey Zhmoginov, Nolan Miller 等ICML 2021 · 被引用 17 次
- Meta-learning families of plasticity rules in recurrent spiking networks using simulation-based inferenceBasile Confavreux, Poornima Ramesh, Pedro J. Gonçalves, Jakob H. Macke 等NeurIPS 2023 · 被引用 18 次
