Brain-Semantoks: Learning Semantic Tokens of Brain Dynamics with a Self-Distilled Foundation Model
Sam Gijsen, Marc-Andre Schulz, Kerstin Ritter
Abstract
The development of foundation models for functional magnetic resonance imaging (fMRI) time series holds significant promise for predicting phenotypes related to disease and cognition. Current models, however, are often trained using a mask-and-reconstruct objective on small brain regions. This focus on low-level information leads to representations that are sensitive to noise and temporal fluctuations, necessitating extensive fine-tuning for downstream tasks. We introduce Brain-Semantoks, a self-supervised framework designed specifically to learn abstract representations of brain dynamics. Its architecture is built on two core innovations: a semantic tokenizer that aggregates noisy regional signals into robust tokens representing functional networks, and a self-distillation objective that enforces representational stability across time. We show that this objective is stabilized through a novel training curriculum, ensuring the model robustly learns meaningful features from low signal-to-noise time series. We demonstrate that learned representations enable strong performance on a variety of downstream tasks even when only using a linear probe. Furthermore, we provide comprehensive scaling analyses indicating more unlabeled data reliably results in out-of-distribution performance gains without domain adaptation.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 200a3759-960f-4bf5-992b-5155a5ccf242Cited by top-tier papers1
Ask how each one uses itBuilds on8
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec et al.NeurIPS 2020 · 9,171 citations
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou et al.ICCV 2021 · 8,921 citations
- Brain Network TransformerXuan Kan, Wei Dai, Hejie Cui, Zilong Zhang et al.NeurIPS 2022 · 272 citations
- BrainLM: A foundation model for brain activity recordingsJosue Ortega Caro, Antonio Henrique de Oliveira Fonseca, Syed Asad Rizvi, Matteo Rosati et al.ICLR 2024 · 109 citations
Related papers
- Large Connectome Model: An fMRI Foundation Model of Brain Connectomes Empowered by Brain-Environment Interaction in Multitask Learning LandscapeZiquan Wei, Tingting Dan, Guorong WuAAAI 2026 · 2 citations
- CalM: A Self-Supervised Foundation Model for Population Dynamics in Calcium Imaging DataXinhong Xu, Yimeng Zhang, Qichen Qian, Yuanlong ZhangICML 2026
- Omni-fMRI: A Universal Atlas-Free fMRI Foundation ModelMo Wang, Wenhao Ye, Junfeng Xia, Junxiang Zhang et al.ICML 2026 · 7 citations
- Unsupervised Representation Learning of Brain Activity via Bridging Voxel Activity and Functional ConnectivityAli Behrouz, Parsa Delavari, Farnoosh HashemiICML 2024 · 8 citations
- RECTOR: Masked Region-Channel-Temporal Modeling for Affective and Cognitive Representation LearningJinhan Liu, Mahsa ShoaranICML 2026
