Neural optimal feedback control with local learning rules
Johannes Friedrich, Siavash Golkar, Shiva Farashahi, Alexander Genkin, Anirvan M. Sengupta, Dmitri B. Chklovskii
摘要
A major problem in motor control is understanding how the brain plans and executes proper movements in the face of delayed and noisy stimuli. A prominent framework for addressing such control problems is Optimal Feedback Control (OFC). OFC generates control actions that optimize behaviorally relevant criteria by integrating noisy sensory stimuli and the predictions of an internal model using the Kalman filter or its extensions. However, a satisfactory neural model of Kalman filtering and control is lacking because existing proposals have the following limitations: not considering the delay of sensory feedback, training in alternating phases, and requiring knowledge of the noise covariance matrices, as well as that of systems dynamics. Moreover, the majority of these studies considered Kalman filtering in isolation, and not jointly with control. To address these shortcomings, we introduce a novel online algorithm which combines adaptive Kalman filtering with a model free control approach (i.e., policy gradient algorithm). We implement this algorithm in a biologically plausible neural network with local synaptic plasticity rules. This network performs system identification and Kalman filtering, without the need for multiple phases with distinct update rules or the knowledge of the noise covariances. It can perform state estimation with delayed sensory feedback, with the help of an internal model. It learns the control policy without requiring any knowledge of the dynamics, thus avoiding the need for weight transport. In this way, our implementation of OFC solves the credit assignment problem needed to produce the appropriate sensory-motor control in the presence of stimulus delay.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Beyond accuracy: generalization properties of bio-plausible temporal credit assignment rulesYuhan Helena Liu, Arna Ghosh, Blake A. Richards, Eric Shea-Brown 等NeurIPS 2022 · 被引用 10 次
- Stochastic Optimal Control and Estimation with Multiplicative and Internal NoiseFrancesco Damiani, Akiyuki Anzai, Jan Drugowitsch, Gregory C. DeAngelis 等NeurIPS 2024 · 被引用 2 次
- Learning Dynamics of RNNs in Closed-Loop EnvironmentsYoav Ger, Omri BarakNeurIPS 2025 · 被引用 2 次
它引用的顶会 Paper2
- Logarithmic Regret Bound in Partially Observable Linear Dynamical SystemsSahin Lale, Kamyar Azizzadenesheli, Babak Hassibi, Anima AnandkumarNeurIPS 2020 · 被引用 106 次
- A Regret Minimization Approach to Iterative Learning ControlNaman Agarwal, Elad Hazan, Anirudha Majumdar, Karan SinghICML 2021 · 被引用 15 次
相关 Paper
- Feedback control guides credit assignment in recurrent neural networksKlara Kaleb, Barbara Feulner, Juan Gallego, Claudia ClopathNeurIPS 2024 · 被引用 5 次
- Credit Assignment in Neural Networks through Deep Feedback ControlAlexander Meulemans, Matilde Tristany Farinha, Javier García Ordóñez, Pau Vilimelis Aceituno 等NeurIPS 2021 · 被引用 61 次
- Real-Time Recurrent Reinforcement LearningJulian Lemmel, Radu GrosuAAAI 2025 · 被引用 8 次
- Minimizing Control for Credit Assignment with Strong FeedbackAlexander Meulemans, Matilde Tristany Farinha, Maria R. Cervera, João Sacramento 等ICML 2022 · 被引用 24 次
- Backprop-Free Reinforcement Learning with Active Neural Generative CodingAlexander G. Ororbia II, Ankur MaliAAAI 2022 · 被引用 23 次
