Stochastic Gradient Descent-Induced Drift of Representation in a Two-Layer Neural Network
Farhad Pashakhanloo, Alexei A. Koulakov
摘要
Representational drift refers to over-time changes in neural activation accompanied by a stable task performance. Despite being observed in the brain and in artificial networks, the mechanisms of drift and its implications are not fully understood. Motivated by recent experimental findings of stimulus-dependent drift in the piriform cortex, we use theory and simulations to study this phenomenon in a two-layer linear feedforward network. Specifically, in a continual online learning scenario, we study the drift induced by the noise inherent in the Stochastic Gradient Descent (SGD). By decomposing the learning dynamics into the normal and tangent spaces of the minimum-loss manifold, we show the former corresponds to a finite variance fluctuation, while the latter could be considered as an effective diffusion process on the manifold. We analytically compute the fluctuation and the diffusion coefficients for the stimuli representations in the hidden layer as functions of network parameters and input distribution. Further, consistent with experiments, we show that the drift rate is slower for a more frequently presented stimulus. Overall, our analysis yields a theoretical framework for better understanding of the drift phenomenon in biological and artificial neural networks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Contribution of task-irrelevant stimuli to drift of neural representationsFarhad PashakhanlooNeurIPS 2025 · 被引用 3 次
- A Model of Place Field Reorganization During Reward MaximizationM. Ganesh Kumar, Blake Bordelon, Jacob A. Zavatone-Veth, Cengiz PehlevanICML 2025
它引用的顶会 Paper1
相关 Paper
- Do Mice Grok? Glimpses of Hidden Progress in Sensory CortexTanishq Kumar, Blake Bordelon, Cengiz Pehlevan, Venkatesh N. Murthy 等ICLR 2025
- SGD vs GD: Rank Deficiency in Linear NetworksAditya Vardhan Varre, Margarita Sagitova, Nicolas FlammarionNeurIPS 2024 · 被引用 5 次
- CogReact: A Reinforced Framework to Model Human Cognitive Reaction Modulated by Dynamic InterventionSonglin Xu, Xinyu ZhangICML 2025
- Dynamical mean-field theory for stochastic gradient descent in Gaussian mixture classificationFrancesca Mignacco, Florent Krzakala, Pierfrancesco Urbani, Lenka ZdeborováNeurIPS 2020 · 被引用 95 次
- SGD with Large Step Sizes Learns Sparse FeaturesMaksym Andriushchenko, Aditya Vardhan Varre, Loucas Pillaud-Vivien, Nicolas FlammarionICML 2023 · 被引用 77 次
