Direct Feedback Alignment Scales to Modern Deep Learning Tasks and Architectures
Julien Launay, Iacopo Poli, François Boniface, Florent Krzakala
摘要
Despite being the workhorse of deep learning, the backpropagation algorithm is no panacea. It enforces sequential layer updates, thus preventing efficient parallelization of the training process. Furthermore, its biological plausibility is being challenged. Alternative schemes have been devised; yet, under the constraint of synaptic asymmetry, none have scaled to modern deep learning tasks and architectures. Here, we challenge this perspective, and study the applicability of Direct Feedback Alignment (DFA) to neural view synthesis, recommender systems, geometric learning, and natural language processing. In contrast with previous studies limited to computer vision tasks, our findings show that it successfully trains a large range of state-of-the-art deep learning architectures, with performance close to fine-tuned backpropagation. When a larger gap between DFA and backpropagation exists, like in Transformers, we attribute this to a need to rethink common practices for large and complex architectures. At variance with common beliefs, our work supports that challenging tasks can be tackled in the absence of weight transport.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper20
- Holomorphic Equilibrium Propagation Computes Exact Gradients Through Finite Size OscillationsAxel Laborieux, Friedemann ZenkeNeurIPS 2022 · 被引用 65 次
- Align, then memorise: the dynamics of learning with feedback alignmentMaria Refinetti, Stéphane d'Ascoli, Ruben Ohana, Sebastian GoldtICML 2021 · 被引用 47 次
- Credit Assignment Through Broadcasting a Global Error VectorDavid G. Clark, L. F. Abbott, SueYeon ChungNeurIPS 2021 · 被引用 29 次
- Forward Learning with Top-Down Feedback: Empirical and Analytical CharacterizationRavi Francesco Srinivasan, Francesca Mignacco, Martino Sorbaro, Maria Refinetti 等ICLR 2024 · 被引用 21 次
- How to Train Your Wide Neural Network Without Backprop: An Input-Weight Alignment PerspectiveAkhilan Boopathy, Ila FieteICML 2022 · 被引用 14 次
它引用的顶会 Paper6
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Implicit Neural Representations with Periodic Activation FunctionsVincent Sitzmann, Julien N. P. Martel, Alexander W. Bergman, David B. Lindell 等NeurIPS 2020 · 被引用 4,008 次
- On the Variance of the Adaptive Learning Rate and BeyondLiyuan Liu, Haoming Jiang, Pengcheng He, Weizhu Chen 等ICLR 2020 · 被引用 2,210 次
- Adaptive Factorization Network: Learning Adaptive-Order Feature InteractionsWeiyu Cheng, Yanyan Shen, Linpeng HuangAAAI 2020 · 被引用 202 次
- Learning to solve the credit assignment problemBenjamin James Lansdell, Prashanth Ravi Prakash, Konrad Paul KördingICLR 2020 · 被引用 60 次
相关 Paper
- Activation Sharing with Asymmetric Paths Solves Weight Transport Problem without Bidirectional ConnectionSunghyeon Woo, Jeongwoo Park, Jiwoo Hong, Dongsuk JeonNeurIPS 2021 · 被引用 3 次
- Learning representations for binary-classification without backpropagationMathias LechnerICLR 2020 · 被引用 6 次
- A Theoretical Framework for Target PropagationAlexander Meulemans, Francesco S. Carzaniga, Johan A. K. Suykens, João Sacramento 等NeurIPS 2020 · 被引用 110 次
- Two Routes to Scalable Credit Assignment without Weight SymmetryDaniel Kunin, Aran Nayebi, Javier Sagastuy-Breña, Surya Ganguli 等ICML 2020 · 被引用 37 次
- Spike-based causal inference for weight alignmentJordan Guerguiev, Konrad P. Körding, Blake A. RichardsICLR 2020 · 被引用 26 次
