Multimodal Scaling Laws for Task & Data-Optimized Models of Visual Cortex
Abdülkadir Gökce, Yingtian Tang, Martin Schrimpf
摘要
Task-optimized neural networks are the leading in-silico models of sensory cortex, yet the field lacks a unified understanding of which modeling choices drive improved brain alignment. Prior NeuroAI work is fragmented across datasets and modalities, making it difficult to determine robust scaling trends. Here, we systematically investigate the scaling laws of model-to-brain alignment across 8 neural datasets (spanning electrophysiology, fMRI, EEG, and MEG) and over 600 models with diverse architectures and pretraining configurations. We report three scaling trends: (1) Pretraining saturation : Alignment improves with pretraining compute and data scale but saturates across all recording modalities. (2) Complementary fine-tuning : Hybrid task & neural data optimization yields consistent improvements in alignment that generalize across datasets and modalities. (3) Mapping scaling : Increasing the number of neural samples to fit model-to-brain mappings yields log-linear gains with the largest impact on alignment. Finally, we propose a novel subject-shared cross-attention mapping which drastically reduces parameter count and improves alignment. Taken together, these results establish multimodal scaling laws that guide resource allocation for next-generation brain models.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper20
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- A ConvNet for the 2020sZhuang Liu, Hanzi Mao, Chao-Yuan Wu, Christoph Feichtenhofer 等CVPR 2022 · 被引用 6,782 次
- Fast is better than free: Revisiting adversarial trainingEric Wong, Leslie Rice, J. Zico KolterICLR 2020 · 被引用 1,352 次
- Perceiver IO: A General Architecture for Structured Inputs & OutputsAndrew Jaegle, Sebastian Borgeaud, Jean-Baptiste Alayrac, Carl Doersch 等ICLR 2022 · 被引用 797 次
相关 Paper
- Brain-tuning Improves Generalizability and Efficiency of Brain Alignment in Speech ModelsOmer Moussa, Mariya TonevaNeurIPS 2025 · 被引用 7 次
- Scaling Laws for Task-Optimized Models of the Primate Visual Ventral StreamAbdülkadir Gökce, Martin SchrimpfICML 2025
- OmniMouse: Scaling properties of multi-modal, multi-task Brain Models on 150B Neural TokensKonstantin Friedrich Willeke, Polina Turishcheva, Alex Gilbert, Goirik Chakrabarty 等ICLR 2026 · 被引用 10 次
- Local Intrinsic Dimension of Representations Predicts Alignment and Generalization in AI Models and Human BrainJunjie Yu, Wenxiao Ma, Chen Wei, Jianyu Zhang 等ICML 2026 · 被引用 2 次
- Alignment between Brains and AI: Evidence for Convergent Evolution across Modalities, Scales and Training TrajectoriesGuobin Shen, Dongcheng Zhao, Yiting Dong, Qian Zhang 等ICML 2026 · 被引用 6 次
