Curl Descent : Non-Gradient Learning Dynamics with Sign-Diverse Plasticity
Hugo Ninou, Jonathan Kadmon, N. Alex Cayco-Gajic
摘要
Gradient-based algorithms are a cornerstone of artificial neural network training, yet it remains unclear whether biological neural networks use similar gradient-based strategies during learning. Experiments often discover a diversity of synaptic plasticity rules, but whether these amount to an approximation to gradient descent is unclear. Here we investigate a previously overlooked possibility: that learning dynamics may include fundamentally non-gradient"curl"-like components while still being able to effectively optimize a loss function. Curl terms naturally emerge in networks with inhibitory-excitatory connectivity or Hebbian/anti-Hebbian plasticity, resulting in learning dynamics that cannot be framed as gradient descent on any objective. To investigate the impact of these curl terms, we analyze feedforward networks within an analytically tractable student-teacher framework, systematically introducing non-gradient dynamics through neurons exhibiting rule-flipped plasticity. Small curl terms preserve the stability of the original solution manifold, resulting in learning dynamics similar to gradient descent. Beyond a critical value, strong curl terms destabilize the solution manifold. Depending on the network architecture, this loss of stability can lead to chaotic learning dynamics that destroy performance. In other cases, the curl terms can counterintuitively speed learning compared to gradient descent by allowing the weight dynamics to escape saddles by temporarily ascending the loss. Our results identify specific architectures capable of supporting robust learning via diverse learning rules, providing an important counterpoint to normative theories of gradient-based learning in neural networks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper9
- Align, then memorise: the dynamics of learning with feedback alignmentMaria Refinetti, Stéphane d'Ascoli, Ruben Ohana, Sebastian GoldtICML 2021 · 被引用 47 次
- Learning to Learn with Feedback and Local PlasticityJack Lindsey, Ashok Litwin-KumarNeurIPS 2020 · 被引用 38 次
- Credit Assignment Through Broadcasting a Global Error VectorDavid G. Clark, L. F. Abbott, SueYeon ChungNeurIPS 2021 · 被引用 29 次
- Constrained Predictive Coding as a Biologically Plausible Model of the Cortical HierarchySiavash Golkar, Tiberiu Tesileanu, Yanis Bahroun, Anirvan M. Sengupta 等NeurIPS 2022 · 被引用 28 次
- Low Tensor Rank Learning of Neural DynamicsArthur Pellegrino, N. Alex Cayco-Gajic, Angus ChadwickNeurIPS 2023 · 被引用 26 次
相关 Paper
- Ubiquity of Emergent Hebbian Dynamics in Regularized LearningDavid Koplow, Tomaso A Poggio, Liu ZiyinICML 2026
- Beyond accuracy: generalization properties of bio-plausible temporal credit assignment rulesYuhan Helena Liu, Arna Ghosh, Blake A. Richards, Eric Shea-Brown 等NeurIPS 2022 · 被引用 10 次
- Spike-timing-dependent Hebbian learning as noisy gradient descentNiklas Dexheimer, Sascha Gaudlitz, Johannes Schmidt-HieberNeurIPS 2025 · 被引用 2 次
- A meta-learning approach to (re)discover plasticity rules that carve a desired function into a neural networkBasile Confavreux, Friedemann Zenke, Everton J. Agnes, Timothy P. Lillicrap 等NeurIPS 2020 · 被引用 40 次
- Tensor decompositions of higher-order correlations by nonlinear Hebbian plasticityGabriel Koch Ocker, Michael A. BuiceNeurIPS 2021 · 被引用 6 次
