Disentanglement with Biological Constraints: A Theory of Functional Cell Types
James C. R. Whittington, Will Dorrell, Surya Ganguli, Timothy Behrens
Abstract
Neurons in the brain are often finely tuned for specific task variables. Moreover, such disentangled representations are highly sought after in machine learning. Here we mathematically prove that simple biological constraints on neurons, namely nonnegativity and energy efficiency in both activity and weights, promote such sought after disentangled representations by enforcing neurons to become selective for single factors of task variation. We demonstrate these constraints lead to disentanglement in a variety of tasks and architectures, including variational autoencoders. We also use this theory to explain why the brain partitions its cells into distinct cell types such as grid and object-vector cells, and also explain when the brain instead entangles representations in response to entangled task factors. Overall, this work provides a mathematical understanding of why single neurons in the brain often represent single human-interpretable factors, and steps towards an understanding task structure shapes the structure of brain representation.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 863cb41f-287f-46d7-8bec-a386e83b5512Cited by top-tier papers23
- Disentanglement via Latent QuantizationKyle Hsu, William Dorrell, James C. R. Whittington, Jiajun Wu et al.NeurIPS 2023 · 54 citations
- Poisson Variational AutoencoderHadi Vafaii, Dekel Galor, Jacob L. YatesNeurIPS 2024 · 18 citations
- Actionable Neural Representations: Grid Cells from Minimal ConstraintsWill Dorrell, Peter E. Latham, Tim E. J. Behrens, James C. R. WhittingtonICLR 2023 · 17 citations
- Binding in hippocampal-entorhinal circuits enables compositionality in cognitive mapsChristopher J. Kymn, Sonia Mazelet, Anthony Thomas, Denis Kleyko et al.NeurIPS 2024 · 14 citations
- Tripod: Three Complementary Inductive Biases for Disentangled Representation LearningKyle Hsu, Jubayer Ibn Hamid, Kaylee Burns, Chelsea Finn et al.ICML 2024 · 13 citations
Builds on8
- VICReg: Variance-Invariance-Covariance Regularization for Self-Supervised LearningAdrien Bardes, Jean Ponce, Yann LeCunICLR 2022 · 1,226 citations
- Relating transformers to models and neural representations of the hippocampal formationJames C. R. Whittington, Joseph Warren, Tim E. J. BehrensICLR 2022 · 110 citations
- Visual Representation Learning Does Not Generalize Strongly Within the Same DomainLukas Schott, Julius von Kügelgen, Frederik Träuble, Peter Vincent Gehler et al.ICLR 2022 · 79 citations
- When Is Unsupervised Disentanglement Possible?Daniella Horan, Eitan Richardson, Yair WeissNeurIPS 2021 · 54 citations
- Demystifying Inductive Biases for (Beta-)VAE Based ArchitecturesDominik Zietlow, Michal Rolínek, Georg MartiusICML 2021 · 24 citations
Related papers
- Range, not Independence, Drives Modularity in Biologically Inspired RepresentationsWill Dorrell, Kyle Hsu, Luke Hollingsworth, Jin Hwa Lee et al.ICLR 2025 · 2 citations
- The role of Disentanglement in GeneralisationMilton Llera Montero, Casimir J. H. Ludwig, Rui Ponte Costa, Gaurav Malhotra et al.ICLR 2021 · 97 citations
- Learning Coherent Representations: A Topological Approach to InterpretabilitySigurd Gaukstad, Melvin Vaupel, Valdemar Kargård Olsen, Erik Hermansen et al.ICML 2026
- Disentangling Representations through Multi-task LearningPantelis Vafidis, Aman Bhargava, Antonio RangelICLR 2025
- Why do networks have inhibitory/negative connections?Qingyang Wang, Michael A. Powell, Ali Geisa, Eric Bridgeford et al.ICCV 2023 · 10 citations
