Global Credit Assignment via Dynamical Criticality
Wentao Wang, Keren Gao, Guozhang Chen
摘要
Efficiently training recurrent neural networks on long sequences remains an open challenge. The standard global paradigm, backpropagation through time (BPTT), suffers from vanishing and exploding gradients and memory costs that scale linearly with sequence length. Conversely, biologically inspired local learning rules are memory-efficient but typically introduce severe bias. To bridge this gap, we introduce Criticality-driven Online Local Alignment (COLA). By leveraging the long-range spatiotemporal correlations inherent to the critical regime, COLA enables a strictly local learning rule to approximate global error propagation, thereby combining online efficiency with gradient descent precision. Theoretically, for a recurrent neural network with hidden units, COLA requires only an auxiliary state and constant activation memory, completely independent of sequence length. Empirically, COLA is competitive with BPTT on standard benchmarks and demonstrates superior robustness on stability-sensitive tasks. Finally, we conduct a rigorous analysis of the approximation error to provide a theoretical foundation for reliable online learning. Code is available at https://github.com/Criticality-Cognitive-Computation-Lab/COLA
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper3
- Training Recurrent Neural Networks via Forward Propagation Through TimeAnil Kag, Venkatesh SaligramaICML 2021 · 被引用 48 次
- Practical Real Time Recurrent Learning with a Sparse ApproximationJacob Menick, Erich Elsen, Utku Evci, Simon Osindero 等ICLR 2021 · 被引用 18 次
- Feedback control guides credit assignment in recurrent neural networksKlara Kaleb, Barbara Feulner, Juan Gallego, Claudia ClopathNeurIPS 2024 · 被引用 5 次
相关 Paper
- Dynamics and Representation Structure of Local Approximations to Gradient-Based Learning in Linear Recurrent Neural NetworksEzekiel Williams, Alexandre Payeur, Guillaume LajoieICML 2026
- RNNs Incrementally Evolving on an Equilibrium Manifold: A Panacea for Vanishing and Exploding Gradients?Anil Kag, Ziming Zhang, Venkatesh SaligramaICLR 2020 · 被引用 51 次
- Training Recurrent Neural Networks Online by Learning Explicit State VariablesSomjit Nath, Vincent Liu, Alan Chan, Xin Li 等ICLR 2020 · 被引用 9 次
- Untangling tradeoffs between recurrence and self-attention in artificial neural networksGiancarlo Kerg, Bhargav Kanuparthi, Anirudh Goyal, Kyle Goyette 等NeurIPS 2020 · 被引用 16 次
- UnICORNN: A recurrent model for learning very long time dependenciesT. Konstantin Rusch, Siddhartha MishraICML 2021 · 被引用 76 次
