Otters: An Energy-Efficient Spiking Transformer via Optical Time-to-First-Spike Encoding
Zhanglu Yan, Jiayi Mao, Qianhui Liu, Fanfan Li, Tao Luo, Gang Pan, Bowen Zhu, Weng-Fai Wong
摘要
Spiking neural networks (SNNs) promise high energy efficiency, particularly with time-to-first-spike (TTFS) encoding, which maximizes sparsity by emitting at most one spike per neuron. However, this energy advantage is often unrealized because inference requires evaluating a temporal decay function and then multiplying the result by the synaptic weights. This paper challenges this costly approach by repurposing a physical hardware `bug', namely, the natural signal decay in optoelectronic devices, as the core computation of TTFS. We fabricated a custom indium oxide optoelectronic synapse that demonstrates how its intrinsic physical decay directly implements the required temporal function. By treating the device's analog output as the fused product of the synaptic weight and temporal decay, optoelectronic synaptic TTFS (named Otters) eliminates these expensive digital operations. To use the Otters' paradigm in complex architectures such as the transformer, which are challenging to train directly due to sparsity, we introduce a novel quantized neural network-to-SNN conversion algorithm. This complete hardware-software co-design enables our model to achieve state-of-the-art accuracy across seven GLUE benchmark datasets and demonstrates a 1.77 improvement in energy efficiency over previous leading SNNs, based on a comprehensive analysis of compute, data movement, and memory access costs using energy measurements from a commercial 22nm process. Our work thus establishes a new paradigm for energy-efficient SNNs that translates fundamental device physics directly into powerful computational primitives. All codes and data are open source.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper9
- TernaryBERT: Distillation-aware Ultra-low Bit BERTWei Zhang, Lu Hou, Yichun Yin, Lifeng Shang 等EMNLP 2020 · 被引用 147 次
- BiT: Robustly Binarized Multi-distilled TransformerZechun Liu, Barlas Oguz, Aasish Pappu, Lin Xiao 等NeurIPS 2022 · 被引用 93 次
- SpikingBERT: Distilling BERT to Train Spiking Language Models Using Implicit DifferentiationMalyaban Bal, Abhronil SenguptaAAAI 2024 · 被引用 78 次
- Temporal-Coded Spiking Neural Networks with Dynamic Firing Threshold: Learning with Event-Driven BackpropagationWenjie Wei, Malu Zhang, Hong Qu, Ammar Belatreche 等ICCV 2023 · 被引用 41 次
- Efficiently Training Time-to-First-Spike Spiking Neural Networks from ScratchKaiwei Che, Wei Fang, Zhengyu Ma, Yifan Huang 等ICML 2026 · 被引用 3 次
相关 Paper
- Training-Free ANN-to-SNN Conversion for High-Performance Spiking TransformersJingya Wang, Xin Deng, Wenjie Wei, Dehao Zhang 等AAAI 2026 · 被引用 1 次
- TTFSFormer: A TTFS-based Lossless Conversion of Spiking TransformerLusen Zhao, Zihan Huang, Jianhao Ding, Zhaofei YuICML 2025
- A time-to-first-spike coding and conversion aware training for energy-efficient deep spiking neural network processor designDongwoo Lew, Kyungchul Lee, Jongsun ParkDAC 2022 · 被引用 18 次
- Parallel Training Time-to-First-Spike Spiking Neural NetworksKaiwei Che, Wei Fang, Peng Xue, Yifan Huang 等AAAI 2026
- Temporal-coded Spiking TransformerQian Sun, Chengzhuo Lu, Wenyu Chen, Wenjie Wei 等ACM MM 2025 · 被引用 1 次
