Gated Linear Networks
Joel Veness, Tor Lattimore, David Budden, Avishkar Bhoopchand, Christopher Mattern, Agnieszka Grabska-Barwinska, Eren Sezener, Jianan Wang, Peter Toth, Simon Schmitt, Marcus Hutter
Abstract
This paper presents a new family of backpropagation-free neural architectures, Gated Linear Networks (GLNs). What distinguishes GLNs from contemporary neural networks is the distributed and local nature of their credit assignment mechanism; each neuron directly predicts the target, forgoing the ability to learn feature representations in favor of rapid online learning. Individual neurons can model nonlinear functions via the use of data-dependent gating in conjunction with online convex optimization. We show that this architecture gives rise to universal learning capabilities in the limit, with effective model capacity increasing as a function of network size in a manner comparable with deep ReLU networks. Furthermore, we demonstrate that the GLN learning mechanism possesses extraordinary resilience to catastrophic forgetting, performing comparably to a MLP with dropout and Elastic Weight Consolidation on standard benchmarks. These desirable theoretical and empirical properties position GLNs as a complementary technique to contemporary offline deep learning methods. * Equal contribution 1 DeepMind.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers5
- MesaNet: Sequence Modeling by Locally Optimal Test-Time TrainingJohannes von Oswald, Nino Scherrer, Seijin Kobayashi, Luca Versari et al.ICLR 2026 · 45 citations
- Scaling Forward Gradient With Local LossesMengye Ren, Simon Kornblith, Renjie Liao, Geoffrey E. HintonICLR 2023 · 12 citations
- The Influence of Learning Rule on Representation Dynamics in Wide Neural NetworksBlake Bordelon, Cengiz PehlevanICLR 2023 · 7 citations
- Geometry-Aware Probabilistic Circuits via Voronoi TessellationsSahil Sidheekh, Sriraam NatarajanICML 2026 · 1 citation
- Deep Networks Learn Features From Local Discontinuities in the Label FunctionPrithaj Banerjee, Harish Guruprasad Ramaswamy, Mahesh Lorik Yadav, Chandra Shekar LakshminarayananICLR 2025
Related papers
- Gaussian Gated Linear NetworksDavid Budden, Adam H. Marblestone, Eren Sezener, Tor Lattimore et al.NeurIPS 2020 · 12 citations
- Globally Gated Deep Linear NetworksQianyi Li, Haim SompolinskyNeurIPS 2022 · 17 citations
- Online Learning in Contextual Bandits using Gated Linear NetworksEren Sezener, Marcus Hutter, David Budden, Jianan Wang et al.NeurIPS 2020 · 10 citations
- A Combinatorial Perspective on Transfer LearningJianan Wang, Eren Sezener, David Budden, Marcus Hutter et al.NeurIPS 2020 · 9 citations
- Make Haste Slowly: A Theory of Emergent Structured Mixed Selectivity in Feature Learning ReLU NetworksDevon Jarvis, Richard Klein, Benjamin Rosman, Andrew M. SaxeICLR 2025
