Attention-Gated Brain Propagation: How the brain can implement reward-based error backpropagation
Isabella Pozzi, Sander M. Bohté, Pieter R. Roelfsema
摘要
Much recent work has focused on biologically plausible variants of supervised learning algorithms. However, there is no teacher in the motor cortex that instructs the motor neurons and learning in the brain depends on reward and punishment. We demonstrate a biologically plausible reinforcement learning scheme for deep networks with an arbitrary number of layers. The network chooses an action by selecting a unit in the output layer and uses feedback connections to assign credit to the units in successively lower layers that are responsible for this action. After the choice, the network receives reinforcement and there is no teacher correcting the errors. We show how the new learning scheme – Attention-Gated Brain Propagation (BrainProp) – is mathematically equivalent to error backpropagation, for one output unit at a time. We demonstrate successful learning of deep fully connected, convolutional and locally connected networks on classical and hard image-classification benchmarks; MNIST, CIFAR10, CIFAR100 andTiny ImageNet. BrainProp achieves an accuracy that is equivalent to that of standard error-backpropagation, and better than state-of-the-art biologically inspired learning schemes. Additionally, the trial-and-error nature of learning is associated with limited additional training time so that BrainProp is a factor of 1-3.5 times slower. Our results thereby provide new insights into how deep learning may be implemented in the brain.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Sparse Spiking Gradient DescentNicolas Perez Nieves, Dan F. M. GoodmanNeurIPS 2021 · 被引用 105 次
- Local plasticity rules can learn deep representations using self-supervised contrastive predictionsBernd Illing, Jean Ventura, Guillaume Bellec, Wulfram GerstnerNeurIPS 2021 · 被引用 99 次
- The least-control principle for local learning at equilibriumAlexander Meulemans, Nicolas Zucchet, Seijin Kobayashi, Johannes von Oswald 等NeurIPS 2022 · 被引用 32 次
- Wiring Up Vision: Minimizing Supervised Synaptic Updates Needed to Produce a Primate Ventral StreamFranziska Geiger, Martin Schrimpf, Tiago Marques, James J. DiCarloICLR 2022 · 被引用 14 次
- Learning by Competition of Self-Interested Reinforcement Learning AgentsStephen ChungAAAI 2022 · 被引用 5 次
相关 Paper
- Hebbian Deep Learning Without FeedbackAdrien Journé, Hector Garcia Rodriguez, Qinghai Guo, Timoleon MoraitisICLR 2023 · 被引用 17 次
- Learning to solve the credit assignment problemBenjamin James Lansdell, Prashanth Ravi Prakash, Konrad Paul KördingICLR 2020 · 被引用 60 次
- Single-phase deep learning in cortico-cortical networksWill Greedy, Heng Wei Zhu, Joseph Pemberton, Jack Mellor 等NeurIPS 2022 · 被引用 62 次
- MAP Propagation Algorithm: Faster Learning with a Team of Reinforcement Learning AgentsStephen ChungNeurIPS 2021 · 被引用 5 次
- Biological credit assignment through dynamic inversion of feedforward networksWilliam F. Podlaski, Christian K. MachensNeurIPS 2020 · 被引用 26 次
