A Model of Place Field Reorganization During Reward Maximization
M. Ganesh Kumar, Blake Bordelon, Jacob A. Zavatone-Veth, Cengiz Pehlevan
摘要
When rodents learn to navigate in a novel environment, a high density of place fields emerges at reward locations, fields elongate against the trajectory, and individual fields change spatial selectivity while demonstrating stable behavior. Why place fields demonstrate these characteristic phenomena during learning remains elusive. We develop a normative framework using a reward maximization objective, whereby the temporal difference (TD) error drives place field reorganization to improve policy learning. Place fields are modeled using Gaussian radial basis functions to represent states in an environment, and directly synapse to an actor-critic for policy learning. Each field's amplitude, center, and width, as well as downstream weights, are updated online at each time step to maximize rewards. We demonstrate that this framework unifies three disparate phenomena observed in navigation experiments. Furthermore, we show that these place field phenomena improve policy convergence when learning to navigate to a single target and relearning multiple new targets. To conclude, we develop a simple normative model that recapitulates several aspects of hippocampal place field learning dynamics and unifies mechanisms to offer testable predictions for future experiments.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper5
- Grounded Language Learning Fast and SlowFelix Hill, Olivier Tieleman, Tamara von Glehn, Nathaniel Wong 等ICLR 2021 · 被引用 85 次
- No Free Lunch from Deep Learning in Neuroscience: A Case Study through Models of the Entorhinal-Hippocampal CircuitRylan Schaeffer, Mikail Khona, Ila FieteNeurIPS 2022 · 被引用 81 次
- Predictive auxiliary objectives in deep RL mimic learning in the brainChing Fang, Kim StachenfeldICLR 2024 · 被引用 16 次
- Stochastic Gradient Descent-Induced Drift of Representation in a Two-Layer Neural NetworkFarhad Pashakhanloo, Alexei A. KoulakovICML 2023 · 被引用 8 次
- Loss Dynamics of Temporal Difference Reinforcement LearningBlake Bordelon, Paul Masset, Henry Kuo, Cengiz PehlevanNeurIPS 2023
相关 Paper
- Emergence of Spatial Representation in an Actor-Critic Agent with Hippocampus-Inspired Sequence GeneratorXiao-Xiong Lin, Yuk Hoi Yiu, Christian LeiboldICLR 2026 · 被引用 2 次
- Time Makes Space: Emergence of Place Fields in Networks Encoding Temporally Continuous Sensory ExperiencesZhaoze Wang, Ronald W. Di Tullio, Spencer Rooke, Vijay BalasubramanianNeurIPS 2024 · 被引用 16 次
- Place Cells as Multi-Scale Position Embeddings: Random Walk Transition Kernels for Path PlanningMinglu Zhao, Dehong Xu, Deqian Kong, Wenhao Zhang 等NeurIPS 2025 · 被引用 1 次
- Trading Place for Space: Increasing Location Resolution Reduces Contextual Capacity in Hippocampal CodesSpencer Rooke, Zhaoze Wang, Ronald W. Di Tullio, Vijay BalasubramanianNeurIPS 2024 · 被引用 4 次
- Learning Place Cell Representations and Context-Dependent RemappingMarkus Pettersen, Frederik Rogge, Mikkel E. LepperødNeurIPS 2024 · 被引用 7 次
