Decoding Dynamic Visual Experience from Calcium Imaging via Cell-Pattern-Aware Pretraining
Sangyoon Bae, Mehdi Azabou, Blake Aaron Richards, Jiook Cha
摘要
Neural recordings exhibit a distinctive form of heterogeneity rooted in differences in cell types, intrinsic circuit dynamics, and stochastic stimulus-response variability that goes beyond ordinary dataset variability, mixing statistically regular neurons with highly stochastic, stimulus-contingent ones within the same dataset. This heterogeneity poses a challenge for self-supervised learning (SSL)-learnable statistical regularity-thereby destabilizing representation learning and limiting reliable scaling. We introduce POYO-CAP (Cell-pattern Aware Pretraining), a biologically grounded hybrid pretraining strategy that first trains with masked reconstruction plus lightweight auxiliary supervision on statistically regular neurons-identified via skewness and kurtosis-and then finetunes on more stochastic populations. On the Allen Brain Observatory dataset, this curriculum yields 12-13% relative improvements over from-scratch training and enables smooth, monotonic scaling with model size, whereas baselines trained on mixed populations plateau or destabilize. By making statistical predictability an explicit data-selection criterion, POYO-CAP turns neural heterogeneity into a scalable learning advantage for robust neural decoding. However, successful SSL relies on statistical regularities in data, as evidenced by masked modeling and sequence prediction in structured domains such as language (
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper9
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec 等NeurIPS 2020 · 被引用 9,171 次
- Deep Learning on a Data Diet: Finding Important Examples Early in TrainingMansheej Paul, Surya Ganguli, Gintare Karolina DziugaiteNeurIPS 2021 · 被引用 806 次
- Masked Language Modeling and the Distributional Hypothesis: Order Word Matters Pre-training for LittleKoustuv Sinha, Robin Jia, Dieuwke Hupkes, Joelle Pineau 等EMNLP 2021 · 被引用 177 次
- The Heavy-Tail Phenomenon in SGDMert Gürbüzbalaban, Umut Simsekli, Lingjiong ZhuICML 2021 · 被引用 165 次
相关 Paper
- Multi-session, multi-task neural decoding from distinct cell-types and brain regionsMehdi Azabou, Krystal Xuejing Pan, Vinam Arora, Ian Jarratt Knight 等ICLR 2025
- The Brain's Bitter Lesson: Scaling Speech Decoding With Self-Supervised LearningDulhan Jayalath, Gilad Landau, Brendan Shillingford, Mark W. Woolrich 等ICML 2025
- Population Transformer: Learning Population-level Representations of Neural ActivityGeeling Chau, Christopher Wang, Sabera J. Talukder, Vighnesh Subramaniam 等ICLR 2025
- Know Thyself by Knowing Others: Learning Neuron Identity from Population ContextVinam Arora, Divyansha Lachi, Ian Jarratt Knight, Mehdi Azabou 等NeurIPS 2025 · 被引用 3 次
- Generalizable, real-time neural decoding with hybrid state-space modelsAvery Hee-Woon Ryoo, Nanda H. Krishna, Ximeng Mao, Mehdi Azabou 等NeurIPS 2025 · 被引用 16 次
