On the Stability and Scalability of Node Perturbation Learning
Naoki Hiratani, Yash Mehta, Timothy P. Lillicrap, Peter E. Latham
摘要
To survive, animals must adapt synaptic weights based on external stimuli and rewards. And they must do so using local, biologically plausible, learning rules – a highly nontrivial constraint. One possible approach is to perturb neural activity (or use intrinsic, ongoing noise to perturb it), determine whether performance increases or decreases, and use that information to adjust the weights. This algorithm – known as node perturbation – has been shown to work on simple problems, but little is known about either its stability or its scalability with respect to network size. We investigate these issues both analytically, in deep linear networks, and numerically, in deep nonlinear ones. We show analytically that in deep linear networks with one hidden layer, both learning time and performance depend very weakly on hidden layer size. However, unlike stochastic gradient descent, when there is model mismatch between the student and teacher networks, node perturbation is always unstable. The instability is triggered by weight diffusion, which eventually leads to very large weights. This instability can be suppressed by weight normalization, at the cost of bias in the learning rule. We confirm numerically that a similar instability, and to a lesser extent scalability, exist in deep nonlinear networks trained on both a motor control task and image classification tasks. Our study highlights the limitations and potential of node perturbation as a biologically plausible learning rule in the brain.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Disentangling and mitigating the impact of task similarity for continual learningNaoki HirataniNeurIPS 2024 · 被引用 21 次
- Towards training digitally-tied analog blocks via hybrid gradient computationTimothy Nest, Maxence ErnoultNeurIPS 2024 · 被引用 6 次
- Zeroth-Order Forward-Only SNN Training Inspiring Neuromorphic On-Chip LearningMingyue Qin, Shuyu Yin, Qinghai Guo, Peilin Liu 等ICML 2026
它引用的顶会 Paper3
- The Impact of Neural Network Overparameterization on Gradient Confusion and Stochastic Gradient DescentKarthik Abinav Sankararaman, Soham De, Zheng Xu, W. Ronny Huang 等ICML 2020 · 被引用 122 次
- Gradientless Descent: High-Dimensional Zeroth-Order OptimizationDaniel Golovin, John Karro, Greg Kochanski, Chansoo Lee 等ICLR 2020 · 被引用 85 次
- Learning to solve the credit assignment problemBenjamin James Lansdell, Prashanth Ravi Prakash, Konrad Paul KördingICLR 2020 · 被引用 60 次
相关 Paper
- Learning by Competition of Self-Interested Reinforcement Learning AgentsStephen ChungAAAI 2022 · 被引用 5 次
- Curl Descent : Non-Gradient Learning Dynamics with Sign-Diverse PlasticityHugo Ninou, Jonathan Kadmon, N. Alex Cayco-GajicNeurIPS 2025 · 被引用 2 次
- Kernelized information bottleneck leads to biologically plausible 3-factor Hebbian learning in deep networksRoman Pogodin, Peter E. LathamNeurIPS 2020 · 被引用 48 次
- Attention-Gated Brain Propagation: How the brain can implement reward-based error backpropagationIsabella Pozzi, Sander M. Bohté, Pieter R. RoelfsemaNeurIPS 2020 · 被引用 38 次
- Two Routes to Scalable Credit Assignment without Weight SymmetryDaniel Kunin, Aran Nayebi, Javier Sagastuy-Breña, Surya Ganguli 等ICML 2020 · 被引用 37 次
