One Forward is Enough for Neural Network Training via Likelihood Ratio Method
Jinyang Jiang, Zeliang Zhang, Chenliang Xu, Zhaofei Yu, Yijie Peng
Abstract
While backpropagation (BP) is the mainstream approach for gradient computation in neural network training, its heavy reliance on the chain rule of differentiation constrains the designing flexibility of network architecture and training pipelines. We avoid the recursive computation in BP and develop a unified likelihood ratio (ULR) method for gradient estimation with just one forward propagation. Not only can ULR be extended to train a wide variety of neural network architectures, but the computation flow in BP can also be rearranged by ULR for better device adaptation. Moreover, we propose several variance reduction techniques to further accelerate the training process. Our experiments offer numerical results across diverse aspects, including various neural network training scenarios, computation flow rearrangement, and fine-tuning of pre-trained models. All findings demonstrate that ULR effectively enhances the flexibility of neural network training by permitting localized module training without compromising the global objective and significantly boosts the network robustness.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext cdbc1085-6e3d-4aab-80cc-0b00f004604aCited by top-tier papers6
- Learning to Transform Dynamically for Better Adversarial TransferabilityRongyi Zhu, Zeliang Zhang, Zhuo Liu, Chenliang Xu et al.CVPR 2024 · 18 citations
- Half-order Fine-Tuning for Diffusion Model: A Recursive Likelihood Ratio OptimizerTao Ren, Zishi Zhang, Jinyang Jiang, Zehao Li et al.ICLR 2026 · 5 citations
- Discover and Mitigate Multiple Biased Subgroups in Image ClassifiersZeliang Zhang, Mingqian Feng, Zhiheng Li, Chenliang XuCVPR 2024 · 5 citations
- Online Pseudo-Zeroth-Order Training of Neuromorphic Spiking Neural NetworksMingqing Xiao, Qingyan Meng, Zongpeng Zhang, Di He et al.ICLR 2026 · 2 citations
- FLOPS: Forward Learning with OPtimal SamplingTao Ren, Zishi Zhang, Jinyang Jiang, Guanghao Li et al.ICLR 2025
Builds on3
- The HSIC Bottleneck: Deep Learning without Back-PropagationKurt Wan-Duo Ma, J. P. Lewis, W. Bastiaan KleijnAAAI 2020 · 180 citations
- Training Spiking Neural Networks with Event-driven BackpropagationYaoyu Zhu, Zhaofei Yu, Wei Fang, Xiaodong Xie et al.NeurIPS 2022 · 57 citations
- TACR-Net: Editing on Deep Video and Voice PortraitsLuchuan Song, Bin Liu, Guojun Yin, Xiaoyi Dong et al.ACM MM 2021 · 20 citations
Related papers
- Efficient Neural Network Training via Forward and Backward Propagation SparsificationXiao Zhou, Weizhong Zhang, Zonghao Chen, Shizhe Diao et al.NeurIPS 2021 · 57 citations
- Accelerated training through iterative gradient propagation along the residual pathErwan Fagnou, Paul Caillon, Blaise Delattre, Alexandre AllauzenICLR 2025
- ADA-GP: Accelerating DNN Training By Adaptive Gradient PredictionVahid Janfaza, Shantanu Mandal, Farabi Mahmud, Abdullah MuzahidMICRO 2023 · 3 citations
- Backpropagation-Free Deep Learning with Recursive Local Representation AlignmentAlexander G. Ororbia II, Ankur Mali, Daniel Kifer, C. Lee GilesAAAI 2023 · 19 citations
- Efficient Backpropagation with Variance Controlled Adaptive SamplingZiteng Wang, Jianfei Chen, Jun ZhuICLR 2024 · 5 citations
