Flexible inference for animal learning rules using neural networks
Yuhan Helena Liu, Victor Geadah, Jonathan W. Pillow
Abstract
Understanding how animals learn is a central challenge in neuroscience, with growing relevance to the development of animal- or human-aligned artificial intelligence. However, existing approaches tend to assume fixed parametric forms for the learning rule (e.g., Q-learning, policy gradient), which may not accurately describe the complex forms of learning employed by animals in realistic settings. Here we address this gap by developing a framework to infer learning rules directly from behavioral data collected during de novo task learning. We assume that animals follow a decision policy parameterized by a generalized linear model (GLM), and we model their learning rule—the mapping from task covariates to per-trial weight updates—using a deep neural network (DNN). This formulation allows flexible, data-driven inference of learning rules while maintaining an interpretable form of the decision policy itself. To capture more complex learning dynamics, we introduce a recurrent neural network (RNN) variant that relaxes the Markovian assumption that learning depends solely on covariates of the current trial, allowing for learning rules that integrate information over multiple trials. Simulations demonstrate that the framework can recover ground-truth learning rules. We applied our DNN and RNN-based methods to a large behavioral dataset from mice learning to perform a sensory decision-making task and found that they outperformed traditional RL learning rules at predicting the learning trajectories of held-out mice. The inferred learning rules exhibited reward-history–dependent learning dynamics, with larger updates following sequences of rewarded trials. Overall, these methods provide a flexible framework for inferring learning rules from behavioral data in de novo learning tasks, setting the stage for improved animal training protocols and the development of behavioral digital twins.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e8dce36d-e87b-495c-aaed-3a2b97bd0fcbBuilds on15
- Gradient Starvation: A Learning Proclivity in Neural NetworksMohammad Pezeshki, Sékou-Oumar Kaba, Yoshua Bengio, Aaron C. Courville et al.NeurIPS 2021 · 378 citations
- Neural Networks as Kernel Learners: The Silent Alignment EffectAlexander B. Atanasov, Blake Bordelon, Cengiz PehlevanICLR 2022 · 110 citations
- Exact learning dynamics of deep linear networks with prior knowledgeLukas Braun, Clémentine C. J. Dominé, James Fitzgerald, Andrew M. SaxeNeurIPS 2022 · 75 citations
- Dynamic Inverse Reinforcement Learning for Characterizing Animal BehaviorZoe Ashwood, Aditi Jha, Jonathan W. PillowNeurIPS 2022 · 50 citations
- A meta-learning approach to (re)discover plasticity rules that carve a desired function into a neural networkBasile Confavreux, Friedemann Zenke, Everton J. Agnes, Timothy P. Lillicrap et al.NeurIPS 2020 · 40 citations
Related papers
- Inferring learning rules from animal decision-makingZoe Ashwood, Nicholas A. Roy, Ji Hyun Bak, Jonathan W. PillowNeurIPS 2020 · 33 citations
- Cognitive Model Discovery via Disentangled RNNsKevin J. Miller, Maria K. Eckstein, Matt M. Botvinick, Zeb Kurth-NelsonNeurIPS 2023 · 39 citations
- Model Based Inference of Synaptic Plasticity RulesYash Mehta, Danil Tyulmankov, Adithya Rajagopalan, Glenn Turner et al.NeurIPS 2024 · 9 citations
- Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal BehaviorsJingyang Ke, Feiyang Wu, Jiyi Wang, Jeffrey Markowitz et al.ICML 2025
- Deep Reinforcement Learning with Time-Scale Invariant MemoryMd Rysul Kabir, James Mochizuki-Freeman, Zoran TiganjAAAI 2025
