Temporal Separation with Entropy Regularization for Knowledge Distillation in Spiking Neural Networks
Kairong Yu, Chengting Yu, Tianqing Zhang, Xiaochen Zhao, Shu Yang, Hongwei Wang, Qiang Zhang, Qi Xu
摘要
Spiking Neural Networks (SNNs), inspired by the human brain, offer significant computational efficiency through discrete spike-based information transfer. Despite their potential to reduce inference energy consumption, a performance gap persists between SNNs and Artificial Neural Networks (ANNs), primarily due to current training methods and inherent model limitations. While recent research has aimed to enhance SNN learning by employing knowledge distillation (KD) from ANN teacher networks, traditional distillation techniques often overlook the distinctive spatiotemporal properties of SNNs, thus failing to fully leverage their advantages. To overcome these challenge, we propose a novel logit distillation method characterized by temporal separation and entropy regularization. This approach improves existing SNN distillation techniques by performing distillation learning on logits across different time steps, rather than merely on aggregated output features. Furthermore, the integration of entropy regularization stabilizes model optimization and further boosts the performance. Extensive experimental results indicate that our method surpasses prior SNN distillation strategies, whether based on logit distillation, feature distillation, or a combination of both. Our project is available at https://github.com/yukairong/TSER .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- TS-SNN: Temporal Shift Module for Spiking Neural NetworksKairong Yu, Tianqing Zhang, Qi Xu, Gang Pan 等ICML 2025
- ReverB-SNN: Reversing Bit of the Weight and Activation for Spiking Neural NetworksYufei Guo, Yuhan Zhang, Jie Zhou, Xiaode Liu 等ICML 2025
- Positive–Unlabeled Reinforcement Learning Distillation for On-Premise Small ModelsZhiqiang Kou, Junyang Chen, Xin-Qiang Cai, Xiaobo Xia 等ICML 2026
- Beyond Linear Processing: Dendritic Bilinear Integration in Spiking Neural NetworksJingyang Ma, Chongming Liu, Songting Li, Douglas ZhouICLR 2026
- Many Eyes, One Mind: Temporal Multi-Perspective and Progressive Distillation for Spiking Neural NetworksKai Sun, Peibo Duan, Yongsheng Huang, Nanxu Gong 等ICLR 2026
它引用的顶会 Paper17
- Deep Residual Learning in Spiking Neural NetworksWei Fang, Zhaofei Yu, Yanqi Chen, Tiejun Huang 等NeurIPS 2021 · 被引用 857 次
- On the Efficacy of Knowledge DistillationJang Hyun Cho, Bharath HariharanICCV 2019 · 被引用 741 次
- Going Deeper With Directly-Trained Larger Spiking Neural NetworksHanle Zheng, Yujie Wu, Lei Deng, Yifan Hu 等AAAI 2021 · 被引用 694 次
- Temporal Efficient Training of Spiking Neural Network via Gradient Re-weightingShikuang Deng, Yuhang Li, Shanghang Zhang, Shi GuICLR 2022 · 被引用 361 次
- Optimal ANN-SNN Conversion for High-accuracy and Ultra-low-latency Spiking Neural NetworksTong Bu, Wei Fang, Jianhao Ding, Penglin Dai 等ICLR 2022 · 被引用 272 次
相关 Paper
- Efficient Logit-based Knowledge Distillation of Deep Spiking Neural Networks for Full-Range Timestep DeploymentChengting Yu, Xiaochen Zhao, Lei Liu, Shu Yang 等ICML 2025
- A Closer Look at Knowledge Distillation in Spiking Neural Network TrainingXu Liu, Na Xia, Jinxing Zhou, Jingyuan Xu 等AAAI 2026
- Bi-Spectrum Distillation: Addressing Spectral Mismatch in ANN-SNN Knowledge TransferYuxuan Zhang, Yuhang Sun, Wen Yao, Yue Deng 等AAAI 2026
- Constructing Deep Spiking Neural Networks from Artificial Neural Networks with Knowledge DistillationQi Xu, Yaxin Li, Jiangrong Shen, Jian K. Liu 等CVPR 2023
- Synergy Between the Strong and the Weak: Spiking Neural Networks are Inherently Self-DistillersYongqi Ding, Lin Zuo, Mengmeng Jing, Kunshan Yang 等NeurIPS 2025 · 被引用 4 次
