Cross-Modal Representational Knowledge Distillation for Enhanced Spike-informed LFP Modeling
Eray Erturk, Saba Hashemi, Maryam M. Shanechi
Abstract
Local field potentials (LFPs) can be routinely recorded alongside spiking activity in intracortical neural experiments, measure a larger complementary spatiotemporal scale of brain activity for scientific inquiry, and can offer practical advantages over spikes, including greater long-term stability, robustness to electrode degradation, and lower power requirements. Despite these advantages, recent neural modeling frameworks have largely focused on spiking activity since LFP signals pose inherent modeling challenges due to their aggregate, population-level nature, often leading to lower predictive power for downstream task variables such as motor behavior. To address this challenge, we introduce a cross-modal knowledge distillation framework that transfers high-fidelity representational knowledge from pretrained multi-session spike transformer models to LFP transformer models. Specifically, we first train a teacher spike model across multiple recording sessions using a masked autoencoding objective with a session-specific neural tokenization strategy. We then align the latent representations of the student LFP model to those of the teacher spike model. Our results show that the Distilled LFP models consistently outperform single- and multi-session LFP baselines in both fully unsupervised and supervised settings, and can generalize to other sessions without additional distillation while maintaining superior performance. These findings demonstrate that cross-modal knowledge distillation is a powerful and scalable approach for leveraging high-performing spike models to develop more accurate LFP models.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9ff0f48c-e052-4884-84b4-69225607dd54Cited by top-tier papers1
Ask how each one uses itBuilds on20
- FlashAttention: Fast and Memory-Efficient Exact Attention with IO-AwarenessTri Dao, Daniel Y. Fu, Stefano Ermon, Atri Rudra et al.NeurIPS 2022 · 5,493 citations
- Contrastive Representation DistillationYonglong Tian, Dilip Krishnan, Phillip IsolaICLR 2020 · 1,305 citations
- A Unified, Scalable Framework for Neural Population DecodingMehdi Azabou, Vinam Arora, Venkataramana Ganesh, Ximeng Mao et al.NeurIPS 2023 · 136 citations
- Neural Data Transformer 2: Multi-context Pretraining for Neural Spiking ActivityJoel Ye, Jennifer L. Collinger, Leila Wehbe, Robert A. GauntNeurIPS 2023 · 100 citations
- MiniLLM: Knowledge Distillation of Large Language ModelsYuxian Gu, Li Dong, Furu Wei, Minlie HuangICLR 2024 · 95 citations
Related papers
- BaRISTA: Brain Scale Informed Spatiotemporal Representation of Human Intracranial Neural ActivityLucine L. Oganesian, Saba Hashemi, Maryam M. ShanechiNeurIPS 2025 · 7 citations
- Self supervised learning for in vivo localization of microelectrode arrays using raw local field potentialTianxiao He, Malhar Patel, Chenyi Li, Anna Maslarova et al.NeurIPS 2025 · 2 citations
- Coupled Transformer Autoencoder for Disentangling Multi-Region Neural Latent DynamicsRam Dyuthi Sristi, Sowmya Manojna Narasimha, Jingya Huang, Alice Despatin et al.ICLR 2026 · 1 citation
- Population Transformer: Learning Population-level Representations of Neural ActivityGeeling Chau, Christopher Wang, Sabera J. Talukder, Vighnesh Subramaniam et al.ICLR 2025
- Mutual Distillation Extracting Spatial-temporal Knowledge for Lightweight Multi-channel Sleep Stage ClassificationZiyu Jia, Haichao Wang, Yucheng Liu, Tianzi JiangKDD 2024 · 5 citations
