Bi-Spectrum Distillation: Addressing Spectral Mismatch in ANN-SNN Knowledge Transfer
Yuxuan Zhang, Yuhang Sun, Wen Yao, Yue Deng, Hongjue Li
Abstract
Knowledge distillation from Artificial Neural Networks (ANNs) to Spiking Neural Networks (SNNs) is a prominent training paradigm. However, its efficacy is fundamentally limited by a spectral mismatch: SNNs, with their intrinsic low-pass filtering characteristics, struggle to learn high-frequency details from their ANN teachers, creating a bottleneck in knowledge transfer at both the feature and logit levels. To address this, we propose Bi-Spectrum Distillation (BSD), a novel framework that mitigates the mismatch from two complementary perspectives. First, at the feature level, our Spectral Residual Distillation (SRD) enhances the student SNN's features with a parameter-efficient, learnable filter that adaptively compensates for high-frequency information loss, which transforms the student's output to better match the teacher's rich spectral target. Second, at the logits level, our Spectral Semantic Distillation (SSD) enhances fine-grained classification by distilling high-frequency components from teacher-ordered logits. Extensive experiments on CIFAR-10/100, ImageNet, and CIFAR10-DVS demonstrate that BSD achieves new state-of-the-art performance across both CNN and Transformer-based SNNs, validating its effectiveness and broad applicability.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 44801e2f-efed-467a-a0c3-626b71ef898fBuilds on18
- Deep Residual Learning in Spiking Neural NetworksWei Fang, Zhaofei Yu, Yanqi Chen, Tiejun Huang et al.NeurIPS 2021 · 857 citations
- Going Deeper With Directly-Trained Larger Spiking Neural NetworksHanle Zheng, Yujie Wu, Lei Deng, Yifan Hu et al.AAAI 2021 · 694 citations
- Temporal Efficient Training of Spiking Neural Network via Gradient Re-weightingShikuang Deng, Yuhang Li, Shanghang Zhang, Shi GuICLR 2022 · 361 citations
- Optimal ANN-SNN Conversion for High-accuracy and Ultra-low-latency Spiking Neural NetworksTong Bu, Wei Fang, Jianhao Ding, Penglin Dai et al.ICLR 2022 · 272 citations
- GLIF: A Unified Gated Leaky Integrate-and-Fire Neuron for Spiking Neural NetworksXingting Yao, Fanrong Li, Zitao Mo, Jian ChengNeurIPS 2022 · 175 citations
Related papers
- Temporal Separation with Entropy Regularization for Knowledge Distillation in Spiking Neural NetworksKairong Yu, Chengting Yu, Tianqing Zhang, Xiaochen Zhao et al.CVPR 2025
- A Closer Look at Knowledge Distillation in Spiking Neural Network TrainingXu Liu, Na Xia, Jinxing Zhou, Jingyuan Xu et al.AAAI 2026
- Constructing Deep Spiking Neural Networks from Artificial Neural Networks with Knowledge DistillationQi Xu, Yaxin Li, Jiangrong Shen, Jian K. Liu et al.CVPR 2023
- Quantized Spike-driven TransformerXuerui Qiu, Malu Zhang, Jieyuan Zhang, Wenjie Wei et al.ICLR 2025
- Many Eyes, One Mind: Temporal Multi-Perspective and Progressive Distillation for Spiking Neural NetworksKai Sun, Peibo Duan, Yongsheng Huang, Nanxu Gong et al.ICLR 2026
