Trajectory-Aware Spiking DiTs Conversion via Membrane Potential Error-Feedback
Haoran Fang, Tianxing Man, Xingchen Li, Wanli Shi, Jinjie Fang, Bin Gu
Abstract
Diffusion Transformers (DiTs) have achieved state-of-the-art generative performance, yet their iterative denoising process remains computationally expensive and energy-intensive. Spiking Neural Networks (SNNs) offer a promising neuromorphic alternative for energy efficiency; however, the non-differentiable nature of spiking neurons makes direct training difficult, positioning ANNto-SNN conversion as a more practical, trainingfree solution. In this paper, we identify a critical challenge unique to converting DiTs: standard fixed-scale spiking neurons fail to accommodate the highly dynamic activation ranges inherent across denoising steps. This mismatch leads to cumulative errors that significantly degrade generation fidelity. To resolve this, we propose a novel conversion framework featuring Multi-Threshold (MT) neurons and a Membrane Potential Error-Feedback (MPEF) mechanism. MT neurons expand the expressive capacity of discrete spikes by employing a multilevel firing strategy. Concurrently, MPEF exploits the temporal correlation between successive denoising steps to recycle residual membrane potential, effectively compensating for information loss and mitigating distribution shifts without retraining. Extensive experiments on ImageNet demonstrate that our framework achieves competitive generative quality with superior energy efficiency, establishing a new performance benchmark for spiking Diffusion Transformers. Our code is available at https://github.com/ JustVelkhana/SpikingDiT.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 7b91deb5-fd0a-4317-ac30-695376c76231Builds on17
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Scalable Diffusion Models with TransformersWilliam Peebles, Saining XieICCV 2023 · 5,568 citations
- Scaling Rectified Flow Transformers for High-Resolution Image SynthesisPatrick Esser, Sumith Kulal, Andreas Blattmann, Rahim Entezari et al.ICML 2024 · 3,620 citations
- PixArt-α: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image SynthesisJunsong Chen, Jincheng Yu, Chongjian Ge, Lewei Yao et al.ICLR 2024 · 831 citations
Related papers
- Efficient ANN-SNN Conversion with Error Compensation LearningChang Liu, Jiangrong Shen, Xuming Ran, Mingkun Xu et al.ICML 2025
- Towards Training-Free and Accurate ANN-to-SNN Conversion via Activation-Aware RedistributionHonglin Cao, Shuai Wang, Zijian Zhou, Ammar Belatreche et al.AAAI 2026
- SpikingIR: A Novel Converted Spiking Neural Network for Efficient Image RestorationYang Ouyang, Zihan Cheng, Xiaotong Luo, Guoqi Li et al.AAAI 2026
- Generalized Threshold Optimization with Harmony Multi-Threshold Neurons for Accurate ANN-to-SNN ConversionWenhan Zhang, Zihan Huang, Tong Bu, Tiejun Huang et al.AAAI 2026
- Adaptive Calibration: A Unified Conversion Framework of Spiking Neural NetworksZiqing Wang, Yuetong Fang, Jiahang Cao, Hongwei Ren et al.AAAI 2025 · 9 citations
