Sparse Distributed Memory is a Continual Learner
Trenton Bricken, Xander Davies, Deepak Singh, Dmitry Krotov, Gabriel Kreiman
Abstract
Continual learning is a problem for artificial neural networks that their biological counterparts are adept at solving. Building on work using Sparse Distributed Memory (SDM) to connect a core neural circuit with the powerful Transformer model, we create a modified Multi-Layered Perceptron (MLP) that is a strong continual learner. We find that every component of our MLP variant translated from biology is necessary for continual learning. Our solution is also free from any memory replay or task information, and introduces novel methods to train sparse networks that may be broadly applicable.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 439cd157-cf22-40bf-8651-bc500c90b175Cited by top-tier papers5
- Efficient Spiking Neural Networks with Sparse Selective Activation for Continual LearningJiangrong Shen, Wenyao Ni, Qi Xu, Huajin TangAAAI 2024 · 42 citations
- Discrete Key-Value BottleneckFrederik Träuble, Anirudh Goyal, Nasim Rahaman, Michael Curtis Mozer et al.ICML 2023 · 25 citations
- Emergence of Sparse Representations from NoiseTrenton Bricken, Rylan Schaeffer, Bruno A. Olshausen, Gabriel KreimanICML 2023 · 15 citations
- Dimensional Collapse in Transformer Attention Outputs: A Challenge for Sparse Dictionary LearningJunxuan Wang, Xuyang Ge, Wentao Shu, Zhengfu He et al.ICML 2026 · 5 citations
- Robust Selective Activation with Randomized Temporal K-Winner-Take-All in Spiking Neural Networks for Continual LearningJiangrong Shen, Liang Zhao, Qi Xu, Yuqi Yang et al.ICLR 2026
Builds on15
- Hash Layers For Large Sparse ModelsStephen Roller, Sainbayar Sukhbaatar, Arthur Szlam, Jason WestonNeurIPS 2021 · 316 citations
- Large Associative Memory Problem in Neurobiology and Machine LearningDmitry Krotov, John J. HopfieldICLR 2021 · 202 citations
- Sparse GPU kernels for deep learningTrevor Gale, Matei Zaharia, Cliff Young, Erich ElsenSC 2020 · 170 citations
- Principled Weight Initialization for HypernetworksOscar Chang, Lampros Flokas, Hod LipsonICLR 2020 · 87 citations
- Universal Hopfield Networks: A General Framework for Single-Shot Associative Memory ModelsBeren Millidge, Tommaso Salvatori, Yuhang Song, Thomas Lukasiewicz et al.ICML 2022 · 72 citations
Related papers
- Distillation-Guided Structural Transfer for Continual Learning Beyond Sparse Distributed MemoryHuiyan Xue, Xuming Ran, Yaxin Li, Qi Xu et al.AAAI 2026 · 2 citations
- Sparse Coding in a Dual Memory System for Lifelong LearningFahad Sarfraz, Elahe Arani, Bahram ZonoozAAAI 2023 · 35 citations
- Meta-Consolidation for Continual LearningK. J. Joseph, Vineeth Nallure BalasubramanianNeurIPS 2020 · 64 citations
- NISPA: Neuro-Inspired Stability-Plasticity Adaptation for Continual Learning in Sparse NetworksMustafa Burak Gurbuz, Constantine DovrolisICML 2022 · 54 citations
- Rapid Learning without Catastrophic Forgetting in the Morris Water MazeRaymond Wang, Jaedong Hwang, Akhilan Boopathy, Ila R. FieteICML 2024 · 2 citations
