Learning to solve the credit assignment problem
Benjamin James Lansdell, Prashanth Ravi Prakash, Konrad Paul Körding
摘要
Backpropagation is driving today's artificial neural networks (ANNs). However, despite extensive research, it remains unclear if the brain implements this algorithm. Among neuroscientists, reinforcement learning (RL) algorithms are often seen as a realistic alternative: neurons can randomly introduce change, and use unspecific feedback signals to observe their effect on the cost and thus approximate their gradient. However, the convergence rate of such learning scales poorly with the number of involved neurons. Here we propose a hybrid learning approach. Each neuron uses an RL-type strategy to learn how to approximate the gradients that backpropagation would provide. We provide proof that our approach converges to the true gradient for certain classes of networks. In both feedforward and convolutional networks, we empirically show that our approach learns to approximate the gradient, and can match or the performance of exact gradient-based learning. Learning feedback weights provides a biologically plausible mechanism of achieving good performance, without the need for precise, pre-specified learning rules.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper16
- Can the Brain Do Backpropagation? - Exact Implementation of Backpropagation in Predictive Coding NetworksYuhang Song, Thomas Lukasiewicz, Zhenghua Xu, Rafal BogaczNeurIPS 2020 · 被引用 117 次
- On Training Implicit ModelsZhengyang Geng, Xin-Yu Zhang, Shaojie Bai, Yisen Wang 等NeurIPS 2021 · 被引用 111 次
- A Theoretical Framework for Target PropagationAlexander Meulemans, Francesco S. Carzaniga, Johan A. K. Suykens, João Sacramento 等NeurIPS 2020 · 被引用 110 次
- Direct Feedback Alignment Scales to Modern Deep Learning Tasks and ArchitecturesJulien Launay, Iacopo Poli, François Boniface, Florent KrzakalaNeurIPS 2020 · 被引用 94 次
- Credit Assignment in Neural Networks through Deep Feedback ControlAlexander Meulemans, Matilde Tristany Farinha, Javier García Ordóñez, Pau Vilimelis Aceituno 等NeurIPS 2021 · 被引用 61 次
相关 Paper
- Attention-Gated Brain Propagation: How the brain can implement reward-based error backpropagationIsabella Pozzi, Sander M. Bohté, Pieter R. RoelfsemaNeurIPS 2020 · 被引用 38 次
- Convergence and Alignment of Gradient Descent with Random Backpropagation WeightsGanlin Song, Ruitu Xu, John LaffertyNeurIPS 2021 · 被引用 4 次
- Spike-based causal inference for weight alignmentJordan Guerguiev, Konrad P. Körding, Blake A. RichardsICLR 2020 · 被引用 26 次
- MAP Propagation Algorithm: Faster Learning with a Team of Reinforcement Learning AgentsStephen ChungNeurIPS 2021 · 被引用 5 次
- Learning by Competition of Self-Interested Reinforcement Learning AgentsStephen ChungAAAI 2022 · 被引用 5 次
