TopoNets: High performing vision and language models with brain-like topography
Mayukh Deb, Mainak Deb, N. Apurva Ratan Murty
Abstract
Independent contributor Neurons in the brain are organized such that nearby cells tend to share similar functions. AI models lack this organization, and past efforts to introduce topography have often led to tradeoffs between topography and task performance. In this work, we present TopoLoss, a new loss function that promotes spatially organized topographic representations in AI models without significantly sacrificing task performance. TopoLoss is highly adaptable and can be seamlessly integrated into the training of leading model architectures. We validate our method on both vision (ResNet-18, ResNet-50, ViT) and language models (GPT-Neo-125M, NanoGPT), collectively TopoNets. TopoNets are the highest performing supervised topographic models to date, exhibiting brain-like properties such as localized feature processing, lower dimensionality, and increased efficiency. TopoNets also predict responses in the brain and replicate the key topographic signatures observed in the brain's visual and language cortices. Together this work establishes a robust and generalizable framework for integrating topography into leading model architectures, advancing the development of high performing models that more closely emulate the computational strategies of the human brain.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 98a4a0c5-2e80-4382-8a0b-d344e16e6da6Cited by top-tier papers3
- Zero-Shot Performance Prediction for Probabilistic Scaling LawsViktoria Schram, Markus Hiller, Daniel Beck, Trevor CohnNeurIPS 2025 · 2 citations
- TDSNNs: Competitive Topographic Deep Spiking Neural Networks for Visual Cortex ModelingDeming Zhou, Yuetong Fang, Zhaorui Wang, Renjing XuAAAI 2026 · 1 citation
- Model-Guided Microstimulation Steers Primate Visual BehaviorJohannes Mehrer, Ben Lonnqvist, Anna Mitola, Paolo Papale et al.ICLR 2026
Builds on4
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Transformer Feed-Forward Layers Build Predictions by Promoting Concepts in the Vocabulary SpaceMor Geva, Avi Caciularu, Kevin Ro Wang, Yoav GoldbergEMNLP 2022 · 92 citations
- Transformer Feed-Forward Layers Are Key-Value MemoriesMor Geva, Roei Schuster, Jonathan Berant, Omer LevyEMNLP 2021 · 33 citations
- FFCV: Accelerating Training by Removing Data BottlenecksGuillaume Leclerc, Andrew Ilyas, Logan Engstrom, Sung Min Park et al.CVPR 2023
Related papers
- Credit-based self organizing maps: training deep topographic networks with minimal performance degradationAmirozhan Dehghani, Xinyu Qian, Asa Farahani, Pouya BashivanICLR 2025
- TopoLM: brain-like spatio-functional organization in a topographic language modelNeil Rathi, Johannes Mehrer, Badr AlKhamissi, Taha Osama A Binhuraib et al.ICLR 2025
- Relating transformers to models and neural representations of the hippocampal formationJames C. R. Whittington, Joseph Warren, Tim E. J. BehrensICLR 2022 · 110 citations
- Neural Language Models are not Born Equal to Fit Brain Data, but Training HelpsAlexandre Pasquiou, Yair Lakretz, John T. Hale, Bertrand Thirion et al.ICML 2022 · 44 citations
- Disentangling the Factors of Convergence between Brains and DINOv3Joséphine Raugel, Marc Szafraniec, Huy V. Vo, Camille Couprie et al.ICLR 2026
