A Theoretical Framework for Target Propagation
Alexander Meulemans, Francesco S. Carzaniga, Johan A. K. Suykens, João Sacramento, Benjamin F. Grewe
摘要
The success of deep learning, a brain-inspired form of AI, has sparked interest in understanding how the brain could similarly learn across multiple layers of neurons. However, the majority of biologically-plausible learning algorithms have not yet reached the performance of backpropagation (BP), nor are they built on strong theoretical foundations. Here, we analyze target propagation (TP), a popular but not yet fully understood alternative to BP, from the standpoint of mathematical optimization. Our theory shows that TP is closely related to Gauss-Newton optimization and thus substantially differs from BP. Furthermore, our analysis reveals a fundamental limitation of difference target propagation (DTP), a well-known variant of TP, in the realistic scenario of non-invertible neural networks. We provide a first solution to this problem through a novel reconstruction loss that improves feedback weight training, while simultaneously introducing architectural flexibility by allowing for direct feedback connections from the output to each hidden layer. Our theory is corroborated by experimental results that show significant improvements in performance and in the alignment of forward weight updates with loss gradients, compared to DTP. Recent work (Lillicrap et al., 2016; Nøkland, 2016) showed that random feedback connections are sufficient to propagate errors and that feedback does not need to adhere to the layer-wise structure of the forward pathway, thereby indicating that weight transport is not strictly necessary for training multilayered neural networks. However, follow-up work (
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper30
- On Training Implicit ModelsZhengyang Geng, Xin-Yu Zhang, Shaojie Bai, Yisen Wang 等NeurIPS 2021 · 被引用 111 次
- Holomorphic Equilibrium Propagation Computes Exact Gradients Through Finite Size OscillationsAxel Laborieux, Friedemann ZenkeNeurIPS 2022 · 被引用 65 次
- Credit Assignment in Neural Networks through Deep Feedback ControlAlexander Meulemans, Matilde Tristany Farinha, Javier García Ordóñez, Pau Vilimelis Aceituno 等NeurIPS 2021 · 被引用 61 次
- Towards Scaling Difference Target Propagation by Learning Backprop TargetsMaxence Ernoult, Fabrice Normandin, Abhinav Moudgil, Sean Spinney 等ICML 2022 · 被引用 49 次
- Neural Networks with Recurrent Generative FeedbackYujia Huang, James Gornet, Sihui Dai, Zhiding Yu 等NeurIPS 2020 · 被引用 48 次
它引用的顶会 Paper4
- Learning to solve the credit assignment problemBenjamin James Lansdell, Prashanth Ravi Prakash, Konrad Paul KördingICLR 2020 · 被引用 60 次
- Two Routes to Scalable Credit Assignment without Weight SymmetryDaniel Kunin, Aran Nayebi, Javier Sagastuy-Breña, Surya Ganguli 等ICML 2020 · 被引用 37 次
- Biological credit assignment through dynamic inversion of feedforward networksWilliam F. Podlaski, Christian K. MachensNeurIPS 2020 · 被引用 26 次
- Spike-based causal inference for weight alignmentJordan Guerguiev, Konrad P. Körding, Blake A. RichardsICLR 2020 · 被引用 26 次
相关 Paper
- Fixed-Weight Difference Target PropagationTatsukichi Shibuya, Nakamasa Inoue, Rei Kawakami, Ikuro SatoAAAI 2023 · 被引用 6 次
- Efficient Target Propagation by Deriving Analytical SolutionYanhao Bao, Tatsukichi Shibuya, Ikuro Sato, Rei Kawakami 等AAAI 2024 · 被引用 2 次
- Activation Sharing with Asymmetric Paths Solves Weight Transport Problem without Bidirectional ConnectionSunghyeon Woo, Jeongwoo Park, Jiwoo Hong, Dongsuk JeonNeurIPS 2021 · 被引用 3 次
- GAIT-prop: A biologically plausible learning rule derived from backpropagation of errorNasir Ahmad, Marcel A. J. van Gerven, Luca AmbrogioniNeurIPS 2020 · 被引用 28 次
- Direct Feedback Alignment Scales to Modern Deep Learning Tasks and ArchitecturesJulien Launay, Iacopo Poli, François Boniface, Florent KrzakalaNeurIPS 2020 · 被引用 94 次
