Identifying Learning Rules From Neural Network Observables
Aran Nayebi, Sanjana Srivastava, Surya Ganguli, Daniel L. K. Yamins
摘要
The brain modifies its synaptic strengths during learning in order to better adapt to its environment. However, the underlying plasticity rules that govern learning are unknown. Many proposals have been suggested, including Hebbian mechanisms, explicit error backpropagation, and a variety of alternatives. It is an open question as to what specific experimental measurements would need to be made to determine whether any given learning rule is operative in a real biological system. In this work, we take a "virtual experimental" approach to this problem. Simulating idealized neuroscience experiments with artificial neural networks, we generate a large-scale dataset of learning trajectories of aggregate statistics measured in a variety of neural network architectures, loss functions, learning rule hyperparameters, and parameter initializations. We then take a discriminative approach, training linear and simple non-linear classifiers to identify learning rules from features based on these observables. We show that different classes of learning rules can be separated solely on the basis of aggregate statistics of the weights, activations, or instantaneous layer-wise activity changes, and that these results generalize to limited access to the trajectory and held-out architectures and learning curricula. We identify the statistics of each observable that are most relevant for rule identification, finding that statistics from network activities across training are more robust to unit undersampling and measurement noise than those obtained from the synaptic strengths. Our results suggest that activation patterns, available from electrophysiological recordings of post-synaptic activities on the order of several hundred units, frequently measured at wider intervals over the course of learning, may provide a good basis on which to identify learning rules.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- Beyond Geometry: Comparing the Temporal Structure of Computation in Neural Circuits with Dynamical Similarity AnalysisMitchell Ostrow, Adam Eisen, Leo Kozachkov, Ila FieteNeurIPS 2023 · 被引用 60 次
- Credit Assignment Through Broadcasting a Global Error VectorDavid G. Clark, L. F. Abbott, SueYeon ChungNeurIPS 2021 · 被引用 29 次
- Distinguishing Learning Rules with Brain Machine InterfacesJacob P. Portes, Christian Schmid, James M. MurrayNeurIPS 2022 · 被引用 12 次
- Beyond accuracy: generalization properties of bio-plausible temporal credit assignment rulesYuhan Helena Liu, Arna Ghosh, Blake A. Richards, Eric Shea-Brown 等NeurIPS 2022 · 被引用 10 次
- Model Based Inference of Synaptic Plasticity RulesYash Mehta, Danil Tyulmankov, Adithya Rajagopalan, Glenn Turner 等NeurIPS 2024 · 被引用 9 次
它引用的顶会 Paper2
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Two Routes to Scalable Credit Assignment without Weight SymmetryDaniel Kunin, Aran Nayebi, Javier Sagastuy-Breña, Surya Ganguli 等ICML 2020 · 被引用 37 次
相关 Paper
- A meta-learning approach to (re)discover plasticity rules that carve a desired function into a neural networkBasile Confavreux, Friedemann Zenke, Everton J. Agnes, Timothy P. Lillicrap 等NeurIPS 2020 · 被引用 40 次
- Memory by accident: a theory of learning as a byproduct of network stabilizationBasile Confavreux, William Dorrell, Nishil Patel, Andrew M. SaxeNeurIPS 2025 · 被引用 2 次
- Characterizing emergent representations in a space of candidate learning rules for deep networksYinan Cao, Christopher Summerfield, Andrew M. SaxeNeurIPS 2020 · 被引用 11 次
- Curl Descent : Non-Gradient Learning Dynamics with Sign-Diverse PlasticityHugo Ninou, Jonathan Kadmon, N. Alex Cayco-GajicNeurIPS 2025 · 被引用 2 次
- Formalizing locality for normative synaptic plasticity modelsColin Bredenberg, Ezekiel Williams, Cristina Savin, Blake A. Richards 等NeurIPS 2023 · 被引用 14 次
