A Model of Place Field Reorganization During Reward Maximization
M. Ganesh Kumar, Blake Bordelon, Jacob A. Zavatone-Veth, Cengiz Pehlevan
Abstract
When rodents learn to navigate in a novel environment, a high density of place fields emerges at reward locations, fields elongate against the trajectory, and individual fields change spatial selectivity while demonstrating stable behavior. Why place fields demonstrate these characteristic phenomena during learning remains elusive. We develop a normative framework using a reward maximization objective, whereby the temporal difference (TD) error drives place field reorganization to improve policy learning. Place fields are modeled using Gaussian radial basis functions to represent states in an environment, and directly synapse to an actor-critic for policy learning. Each field's amplitude, center, and width, as well as downstream weights, are updated online at each time step to maximize rewards. We demonstrate that this framework unifies three disparate phenomena observed in navigation experiments. Furthermore, we show that these place field phenomena improve policy convergence when learning to navigate to a single target and relearning multiple new targets. To conclude, we develop a simple normative model that recapitulates several aspects of hippocampal place field learning dynamics and unifies mechanisms to offer testable predictions for future experiments.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 24e35a36-998e-4237-a436-d00a8a87d939Cited by top-tier papers1
Ask how each one uses itBuilds on5
- Grounded Language Learning Fast and SlowFelix Hill, Olivier Tieleman, Tamara von Glehn, Nathaniel Wong et al.ICLR 2021 · 85 citations
- No Free Lunch from Deep Learning in Neuroscience: A Case Study through Models of the Entorhinal-Hippocampal CircuitRylan Schaeffer, Mikail Khona, Ila FieteNeurIPS 2022 · 81 citations
- Predictive auxiliary objectives in deep RL mimic learning in the brainChing Fang, Kim StachenfeldICLR 2024 · 16 citations
- Stochastic Gradient Descent-Induced Drift of Representation in a Two-Layer Neural NetworkFarhad Pashakhanloo, Alexei A. KoulakovICML 2023 · 8 citations
- Loss Dynamics of Temporal Difference Reinforcement LearningBlake Bordelon, Paul Masset, Henry Kuo, Cengiz PehlevanNeurIPS 2023
Related papers
- Emergence of Spatial Representation in an Actor-Critic Agent with Hippocampus-Inspired Sequence GeneratorXiao-Xiong Lin, Yuk Hoi Yiu, Christian LeiboldICLR 2026 · 2 citations
- Time Makes Space: Emergence of Place Fields in Networks Encoding Temporally Continuous Sensory ExperiencesZhaoze Wang, Ronald W. Di Tullio, Spencer Rooke, Vijay BalasubramanianNeurIPS 2024 · 16 citations
- Place Cells as Multi-Scale Position Embeddings: Random Walk Transition Kernels for Path PlanningMinglu Zhao, Dehong Xu, Deqian Kong, Wenhao Zhang et al.NeurIPS 2025 · 1 citation
- Trading Place for Space: Increasing Location Resolution Reduces Contextual Capacity in Hippocampal CodesSpencer Rooke, Zhaoze Wang, Ronald W. Di Tullio, Vijay BalasubramanianNeurIPS 2024 · 4 citations
- Learning Place Cell Representations and Context-Dependent RemappingMarkus Pettersen, Frederik Rogge, Mikkel E. LepperødNeurIPS 2024 · 7 citations
