Contribution of task-irrelevant stimuli to drift of neural representations
Farhad Pashakhanloo
Abstract
Biological and artificial learners are inherently exposed to a stream of data and experience throughout their lifetimes and must constantly adapt to, learn from, or selectively ignore the ongoing input. Recent findings reveal that, even when the performance remains stable, the underlying neural representations can change gradually over time, a phenomenon known as representational drift. Studying the different sources of data and noise that may contribute to drift is essential for understanding lifelong learning in neural systems. However, a systematic study of drift across architectures and learning rules, and the connection to task, are missing. Here, in an online learning setup, we characterize drift as a function of data distribution, and specifically show that the learning noise induced by task-irrelevant stimuli, which the agent learns to ignore in a given context, can create long-term drift in the representation of task-relevant stimuli. Using theory and simulations, we demonstrate this phenomenon both in Hebbian-based learning -- Oja's rule and Similarity Matching -- and in stochastic gradient descent applied to autoencoders and a supervised two-layer network. We consistently observe that the drift rate increases with the variance and the dimension of the data in the task-irrelevant subspace. We further show that this yields different qualitative predictions for the geometry and dimension-dependency of drift than those arising from Gaussian synaptic noise. Overall, our study links the structure of stimuli, task, and learning rule to representational drift and could pave the way for using drift as a signal for uncovering underlying computation in the brain.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on3
- What Happens after SGD Reaches Zero Loss? --A Mathematical FrameworkZhiyuan Li, Tianhao Wang, Sanjeev AroraICLR 2022 · 121 citations
- Gradient Descent on Neural Networks Typically Occurs at the Edge of StabilityJeremy Cohen, Simran Kaur, Yuanzhi Li, J. Zico Kolter et al.ICLR 2021 · 22 citations
- Stochastic Gradient Descent-Induced Drift of Representation in a Two-Layer Neural NetworkFarhad Pashakhanloo, Alexei A. KoulakovICML 2023 · 8 citations
Related papers
- Characterizing emergent representations in a space of candidate learning rules for deep networksYinan Cao, Christopher Summerfield, Andrew M. SaxeNeurIPS 2020 · 11 citations
- Deep Reinforcement Learning amidst Continual Structured Non-StationarityAnnie Xie, James Harrison, Chelsea FinnICML 2021 · 43 citations
- Dynamics of Supervised and Reinforcement Learning in the Non-Linear PerceptronChristian Schmid, James M. MurrayNeurIPS 2024
- Learning Representations on the Unit Sphere: Investigating Angular Gaussian and Von Mises-Fisher Distributions for Online Continual LearningNicolas Michel, Giovanni Chierchia, Romain Negrel, Jean-François BercherAAAI 2024 · 14 citations
- A meta-learning approach to (re)discover plasticity rules that carve a desired function into a neural networkBasile Confavreux, Friedemann Zenke, Everton J. Agnes, Timothy P. Lillicrap et al.NeurIPS 2020 · 40 citations
