Can Forward Gradient Match Backpropagation?
Louis Fournier, Stéphane Rivaud, Eugene Belilovsky, Michael Eickenberg, Edouard Oyallon
Abstract
Forward Gradients - the idea of using directional derivatives in forward differentiation mode - have recently been shown to be utilizable for neural network training while avoiding problems generally associated with backpropagation gradient computation, such as locking and memorization requirements. The cost is the requirement to guess the step direction, which is hard in high dimensions. While current solutions rely on weighted averages over isotropic guess vector distributions, we propose to strongly bias our gradient guesses in directions that are much more promising, such as feedback obtained from small, local auxiliary networks. For a standard computer vision neural network, we conduct a rigorous study systematically covering a variety of combinations of gradient targets and gradient guesses, including those previously presented in the literature. We find that using gradients obtained from a local loss as a candidate direction drastically improves on random noise in Forward Gradient methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d1bc978a-a0f6-4707-926d-6539a03307bdCited by top-tier papers10
- Scaling Supervised Local Learning with Augmented Auxiliary NetworksChenxiang Ma, Jibin Wu, Chenyang Si, Kay Chen TanICLR 2024 · 9 citations
- Learning long range dependencies through time reversal symmetry breakingGuillaume Pourcel, Maxence ErnoultNeurIPS 2025 · 9 citations
- Towards training digitally-tied analog blocks via hybrid gradient computationTimothy Nest, Maxence ErnoultNeurIPS 2024 · 6 citations
- Stepping Forward on the Last MileChen Feng, Jay Zhuo, Parker Zhang, Ramchalam Kinattinkara Ramakrishnan et al.NeurIPS 2024 · 4 citations
- Accelerating Legacy Numerical Solvers by Non-intrusive Gradient-based Meta-solvingSohei Arisaka, Qianxiao LiICML 2024 · 1 citation
Builds on4
- Revisiting Locally Supervised Learning: an Alternative to End-to-end TrainingYulin Wang, Zanlin Ni, Shiji Song, Le Yang et al.ICLR 2021 · 99 citations
- Align, then memorise: the dynamics of learning with feedback alignmentMaria Refinetti, Stéphane d'Ascoli, Ruben Ohana, Sebastian GoldtICML 2021 · 47 citations
- Learning by Directional Gradient DescentDavid Silver, Anirudh Goyal, Ivo Danihelka, Matteo Hessel et al.ICLR 2022 · 44 citations
- Scaling Forward Gradient With Local LossesMengye Ren, Simon Kornblith, Renjie Liao, Geoffrey E. HintonICLR 2023 · 12 citations
Related papers
- Minimizing Control for Credit Assignment with Strong FeedbackAlexander Meulemans, Matilde Tristany Farinha, Maria R. Cervera, João Sacramento et al.ICML 2022 · 24 citations
- Neural-Guided RANSAC: Learning Where to Sample Model HypothesesEric Brachmann, Carsten RotherICCV 2019 · 282 citations
- GAIT-prop: A biologically plausible learning rule derived from backpropagation of errorNasir Ahmad, Marcel A. J. van Gerven, Luca AmbrogioniNeurIPS 2020 · 28 citations
- The HSIC Bottleneck: Deep Learning without Back-PropagationKurt Wan-Duo Ma, J. P. Lewis, W. Bastiaan KleijnAAAI 2020 · 180 citations
- Error-driven Input Modulation: Solving the Credit Assignment Problem without a Backward PassGiorgia Dellaferrera, Gabriel KreimanICML 2022 · 80 citations
