Cross-Subject Modeling for Widefield Calcium Imaging via Atlas-Aligned Spatiotemporal Tokenization
Mohammad Hosseini, Eray Erturk, Saba Hashemi, Maryam Shanechi
Abstract
Large-scale, multi-subject widefield calcium imaging provides unprecedented access to brain-wide cortical dynamics. However, the high dimensionality, complex spatiotemporal structure, and substantial task-irrelevant activity in widefield recordings have largely restricted modeling efforts to single-session analyses, limiting scalability and generalization. While multi-subject pretrained models have been explored for some neural modalities, multi-subject models for widefield calcium imaging have not yet been demonstrated; further, subject-invariant zero-shot behavior decoding remains elusive for multi-subject models across neural modalities more broadly. As a first step toward foundation modeling of widefield data, we introduce WiCAT, a multi-subject model that leverages self-supervised pretraining to both outperform single-session models and enable zero-shot behavior decoding on unseen subjects. WiCAT introduces an atlas-grounded tokenization scheme without session-specific components and learns globally shared spatiotemporal representations. Across multiple widefield datasets, the pretrained model supports lightweight downstream decoding, transfers across subjects, tasks, and datasets, and outperforms baseline models. Notably, the model also achieves robust zero-shot continuous behavior decoding and left-out brain region reconstruction on unseen subjects.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on20
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- FlashAttention: Fast and Memory-Efficient Exact Attention with IO-AwarenessTri Dao, Daniel Y. Fu, Stefano Ermon, Atri Rudra et al.NeurIPS 2022 · 5,493 citations
- A Unified, Scalable Framework for Neural Population DecodingMehdi Azabou, Vinam Arora, Venkataramana Ganesh, Ximeng Mao et al.NeurIPS 2023 · 136 citations
- Neural Data Transformer 2: Multi-context Pretraining for Neural Spiking ActivityJoel Ye, Jennifer L. Collinger, Leila Wehbe, Robert A. GauntNeurIPS 2023 · 100 citations
- Brain-JEPA: Brain Dynamics Foundation Model with Gradient Positioning and Spatiotemporal MaskingZijian Dong, Ruilin Li, Yilei Wu, Thuan Tinh Nguyen et al.NeurIPS 2024 · 96 citations
Related papers
- CalM: A Self-Supervised Foundation Model for Population Dynamics in Calcium Imaging DataXinhong Xu, Yimeng Zhang, Qichen Qian, Yuanlong ZhangICML 2026
- Dynamical Modeling of Behaviorally Relevant Spatiotemporal Patterns in Neural Imaging DataSayed Mohammad Hosseini, Maryam ShanechiICML 2025
- Multi-session, multi-task neural decoding from distinct cell-types and brain regionsMehdi Azabou, Krystal Xuejing Pan, Vinam Arora, Ian Jarratt Knight et al.ICLR 2025
- BaRISTA: Brain Scale Informed Spatiotemporal Representation of Human Intracranial Neural ActivityLucine L. Oganesian, Saba Hashemi, Maryam M. ShanechiNeurIPS 2025 · 7 citations
- Self supervised learning for in vivo localization of microelectrode arrays using raw local field potentialTianxiao He, Malhar Patel, Chenyi Li, Anna Maslarova et al.NeurIPS 2025 · 2 citations
