BrainLM: A foundation model for brain activity recordings
Josue Ortega Caro, Antonio Henrique de Oliveira Fonseca, Syed Asad Rizvi, Matteo Rosati, Christopher L. Averill, James L. Cross, Prateek Mittal, Emanuele Zappala, Rahul Madhav Dhodapkar, Chadi Abdallah, David van Dijk
摘要
We introduce the Brain Language Model (BrainLM), a foundation model for brain activity dynamics trained on 6,700 hours of fMRI recordings. Utilizing self-supervised masked-prediction training, BrainLM demonstrates proficiency in both fine-tuning and zero-shot inference tasks. Fine-tuning allows for the accurate prediction of clinical variables like age, anxiety, and PTSD as well as forecasting of future brain states. Critically, the model generalizes well to entirely new external cohorts not seen during training. In zero-shot inference mode, BrainLM can identify intrinsic functional networks directly from raw fMRI data without any network-based supervision during training. The model also generates interpretable latent representations that reveal relationships between brain activity patterns and cognitive states. Overall, BrainLM offers a versatile and interpretable framework for elucidating the complex spatiotemporal dynamics of human brain activity. It serves as a powerful "lens" through which massive repositories of fMRI data can be analyzed in new ways, enabling more effective interpretation and utilization at scale. The work demonstrates the potential of foundation models to advance computational neuroscience research.
Prior work has explored various machine-learning techniques for analyzing fMRI recordings. Earlier approaches focused on decoding cognitive states from activity patterns. Methods like SVM and neural networks were trained in a supervised fashion to classify fMRI data into stimulus categories or regress against variables of interest (Horikawa & Kamitani, 2017;Hoefle et al., 2018;Beliy et al., 2019). However, these models learn representations tailored to specific tasks and struggle to generalize.
Recent work has aimed to obtain more transferable fMRI encodings without task-specific constraints. Techniques include training autoencoders to reconstruct recordings, learning to map recordings to a lower-dimensional space
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper24
- Brain-JEPA: Brain Dynamics Foundation Model with Gradient Positioning and Spatiotemporal MaskingZijian Dong, Ruilin Li, Yilei Wu, Thuan Tinh Nguyen 等NeurIPS 2024 · 被引用 96 次
- Brain Harmony: A Multimodal Foundation Model Unifying Morphology and Function into 1D TokensZijian Dong, Ruilin Li, Joanna Su Xian Chong, Niousha Dehestani 等NeurIPS 2025 · 被引用 24 次
- A Generalist Intracortical Motor DecoderJoel Ye, Fabio Rizzoglio, Xuan Ma, Adam Smoulder 等NeurIPS 2025 · 被引用 21 次
- Context parroting: A simple but tough-to-beat baseline for foundation models in scientific machine learningYuanzhao Zhang, William GilpinICLR 2026 · 被引用 16 次
- Multi-Modal View Enhanced Large Vision Models for Long-Term Time Series ForecastingChengAo Shen, Wenchao Yu, Ziming Zhao, Dongjin Song 等NeurIPS 2025 · 被引用 14 次
它引用的顶会 Paper7
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Cinematic Mindscapes: High-quality Video Reconstruction from Brain ActivityZijiao Chen, Jiaxin Qing, Juan Helen ZhouNeurIPS 2023 · 被引用 109 次
- Neural Language Models are not Born Equal to Fit Brain Data, but Training HelpsAlexandre Pasquiou, Yair Lakretz, John T. Hale, Bertrand Thirion 等ICML 2022 · 被引用 44 次
- Masked Autoencoders Are Scalable Vision LearnersKaiming He, Xinlei Chen, Saining Xie, Yanghao Li 等CVPR 2022
相关 Paper
- fMRI-LM: Towards a Universal Foundation Model for Language-Aligned fMRI UnderstandingYuxiang Wei, Yanteng Zhang, Xi Xiao, Chengxuan Qian 等CVPR 2026 · 被引用 11 次
- Scaling Vision Transformers for Functional MRI with Flat MapsConnor Lane, Mihir Tripathy, Leema K Murali, Ratna Grandhi 等ICML 2026 · 被引用 3 次
- Large Connectome Model: An fMRI Foundation Model of Brain Connectomes Empowered by Brain-Environment Interaction in Multitask Learning LandscapeZiquan Wei, Tingting Dan, Guorong WuAAAI 2026 · 被引用 2 次
- A Brain Graph Foundation Model: Pre-Training and Prompt-Tuning across Broad Atlases and DisordersXinxu Wei, kanhao zhao, Yong Jiao, Lifang He 等ICLR 2026 · 被引用 6 次
- Brain-tuning Improves Generalizability and Efficiency of Brain Alignment in Speech ModelsOmer Moussa, Mariya TonevaNeurIPS 2025 · 被引用 7 次
