Brain Harmony: A Multimodal Foundation Model Unifying Morphology and Function into 1D Tokens
Zijian Dong, Ruilin Li, Joanna Su Xian Chong, Niousha Dehestani, Yinghui Teng, Yi Lin, Zhizhou Li, Yichi Zhang, Yapei Xie, Leon Qi Rong Ooi, B. T. Thomas Yeo, Juan Helen Zhou
摘要
We present Brain Harmony (BrainHarmonix), the first multimodal brain foundation model that unifies structural morphology and functional dynamics into compact 1D token representations. The model was pretrained on two of the largest neuroimaging datasets to date, encompassing 64,594 T1-weighted structural MRI 3D volumes ( 14 million images) and 70,933 functional MRI (fMRI) time series. BrainHarmonix is grounded in two foundational neuroscience principles: structure complements function - structural and functional modalities offer distinct yet synergistic insights into brain organization; function follows structure - brain functional dynamics are shaped by cortical morphology. The modular pretraining process involves single-modality training with geometric pre-alignment followed by modality fusion through shared brain hub tokens. Notably, our dynamics encoder uniquely handles fMRI time series with heterogeneous repetition times (TRs), addressing a major limitation in existing models. BrainHarmonix is also the first to deeply compress high-dimensional neuroimaging signals into unified, continuous 1D tokens, forming a compact latent space of the human brain. BrainHarmonix achieves strong generalization across diverse downstream tasks, including neurodevelopmental and neurodegenerative disorder classification and cognition prediction - consistently outperforming previous approaches. Our models - pretrained on 8 H100 GPUs - aim to catalyze a new era of AI-driven neuroscience powered by large-scale multimodal neuroimaging.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Omni-fMRI: A Universal Atlas-Free fMRI Foundation ModelMo Wang, Wenhao Ye, Junfeng Xia, Junxiang Zhang 等ICML 2026 · 被引用 7 次
- Scaling Vision Transformers for Functional MRI with Flat MapsConnor Lane, Mihir Tripathy, Leema K Murali, Ratna Grandhi 等ICML 2026 · 被引用 3 次
- Stochastic Optimal Control for Continuous-Time fMRI Representation LearningJoonhyeong Park, Byoungwoo Park, Chang-Bae Bang, Jungwon Choi 等ICLR 2026 · 被引用 2 次
它引用的顶会 Paper10
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- FlashAttention: Fast and Memory-Efficient Exact Attention with IO-AwarenessTri Dao, Daniel Y. Fu, Stefano Ermon, Atri Rudra 等NeurIPS 2022 · 被引用 5,493 次
- FlashAttention-2: Faster Attention with Better Parallelism and Work PartitioningTri DaoICLR 2024 · 被引用 2,600 次
- Brain Network TransformerXuan Kan, Wei Dai, Hejie Cui, Zilong Zhang 等NeurIPS 2022 · 被引用 272 次
- BrainLM: A foundation model for brain activity recordingsJosue Ortega Caro, Antonio Henrique de Oliveira Fonseca, Syed Asad Rizvi, Matteo Rosati 等ICLR 2024 · 被引用 109 次
相关 Paper
- Large Connectome Model: An fMRI Foundation Model of Brain Connectomes Empowered by Brain-Environment Interaction in Multitask Learning LandscapeZiquan Wei, Tingting Dan, Guorong WuAAAI 2026 · 被引用 2 次
- BrainMoE: Cognition Joint Embedding via Mixture-of-Expert Towards Robust Brain Foundation ModelZiquan Wei, Tingting Dan, Tianlong Chen, Guorong WuNeurIPS 2025 · 被引用 2 次
- A Brain Graph Foundation Model: Pre-Training and Prompt-Tuning across Broad Atlases and DisordersXinxu Wei, kanhao zhao, Yong Jiao, Lifang He 等ICLR 2026 · 被引用 6 次
- Can Natural Image Autoencoders Compactly Tokenize fMRI Volumes for Long-Range Dynamics Modeling?Peter Yongho Kim, Juhyeon Park, Jungwoo Park, Jubin Choi 等CVPR 2026 · 被引用 1 次
- BrainOmni: A Brain Foundation Model for Unified EEG and MEG SignalsQinfan Xiao, Ziyun Cui, Chi Zhang, Siqi Chen 等NeurIPS 2025 · 被引用 38 次
