LAS: Loss-less ANN-SNN Conversion for Fully Spike-Driven Large Language Models
Long Chen, Xiaotian Song, Yanan Sun
Abstract
Spiking Large Language Models (LLMs) have emerged as an energy-efficient alternative to conventional LLMs through their event-driven computation. To effectively obtain spiking LLMs, researchers develop different ANN-to-SNN conversion methods by leveraging pre-trained ANN parameters while inheriting the energy efficiency of SNN. However, existing conversion methods struggle with extreme activation outliers and incompatible nonlinear operations of ANN-based LLMs. To address this, we propose a loss-less ANN-SNN conversion for fully spike-driven LLMs, termed LAS. Specifically, LAS introduces two novel neurons to convert the activation outlier and nonlinear operation of ANN-based LLMs. Moreover, LAS tailors the spike-equivalent Transformer components for spiking LLMs, which can ensure full spiking conversion without any loss of performance. Experimental results on six language models and two vision-language models demonstrate that LAS achieves loss-less conversion. Notably, on OPT-66B, LAS even improves the accuracy of 2% on the WSC task. In addition, the parameter and ablation studies further verify the effectiveness of LAS. 1
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0b1350bc-fe07-4ac2-9996-40068adea310Cited by top-tier papers2
- Plug-and-Play Spiking Operators: Breaking the Nonlinearity Bottleneck in Spiking TransformersXinzhe Yuan, Xiang Peng, Bin Gu, Huan XiongICML 2026
- Distribution-Aware Multi-Granularity Phase Coding: Towards Lower Conversion Error for Spike-Driven Large Language ModelsHanyuan Zheng, Haozhen Zhang, Tianshuo Chen, Zhaogeng Liu et al.ICLR 2026
Builds on15
- Multitask Prompted Training Enables Zero-Shot Task GeneralizationVictor Sanh, Albert Webson, Colin Raffel, Stephen H. Bach et al.ICLR 2022 · 1,976 citations
- Pythia: A Suite for Analyzing Large Language Models Across Training and ScalingStella Biderman, Hailey Schoelkopf, Quentin Gregory Anthony, Herbie Bradley et al.ICML 2023 · 1,822 citations
- Optimal ANN-SNN Conversion for High-accuracy and Ultra-low-latency Spiking Neural NetworksTong Bu, Wei Fang, Jianhao Ding, Penglin Dai et al.ICLR 2022 · 272 citations
- Optimal Conversion of Conventional Artificial Neural Networks to Spiking Neural NetworksShikuang Deng, Shi GuICLR 2021 · 100 citations
- Reducing ANN-SNN Conversion Error through Residual Membrane PotentialZecheng Hao, Tong Bu, Jianhao Ding, Tiejun Huang et al.AAAI 2023 · 85 citations
Related papers
- Training-Free ANN-to-SNN Conversion for High-Performance Spiking TransformersJingya Wang, Xin Deng, Wenjie Wei, Dehao Zhang et al.AAAI 2026 · 1 citation
- Towards Training-Free and Accurate ANN-to-SNN Conversion via Activation-Aware RedistributionHonglin Cao, Shuai Wang, Zijian Zhou, Ammar Belatreche et al.AAAI 2026
- SpikedAttention: Training-Free and Fully Spike-Driven Transformer-to-SNN Conversion with Winner-Oriented Spike Shift for Softmax OperationSangwoo Hwang, Seunghyun Lee, Dahoon Park, Donghun Lee et al.NeurIPS 2024 · 23 citations
- Towards High-performance Spiking Transformers from ANN to SNN ConversionZihan Huang, Xinyu Shi, Zecheng Hao, Tong Bu et al.ACM MM 2024 · 17 citations
- SpikeZIP-TF: Conversion is All You Need for Transformer-based SNNKang You, Zekai Xu, Chen Nie, Zhijie Deng et al.ICML 2024 · 20 citations
