Hybrid Spiking Vision Transformer for Object Detection with Event Cameras
Qi Xu, Jie Deng, Jiangrong Shen, Biwu Chen, Huajin Tang, Gang Pan
Abstract
Event-based object detection has attracted increasing attention for its high temporal resolution, wide dynamic range, and asynchronous address-event representation. Leveraging these advantages, spiking neural networks (SNNs) have emerged as a promising approach, offering low energy consumption and rich spatiotemporal dynamics. To further enhance the performance of event-based object detection, this study proposes a novel hybrid spike vision Transformer (HsVT) model. The HsVT model integrates a spatial feature extraction module to capture local and global features, and a temporal feature extraction module to model time dependencies and long-term patterns in event sequences. This combination enables HsVT to capture spatiotemporal features, improving its capability in handling complex event-based object detection tasks. To support research in this area, we developed the Fall Detection dataset as a benchmark for event-based object detection tasks. The Fall DVS detection dataset protects facial privacy and reduces memory usage thanks to its event-based representation. Experimental results demonstrate that HsVT outperforms existing SNN methods and achieves competitive performance compared to ANN-based models, with fewer parameters and lower energy consumption.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 065da5bf-c75a-40b7-b21e-3a4d2a8416fbCited by top-tier papers5
- FlashCap: Millisecond-Accurate Human Motion Capture via Flashing LEDs and Event-Based VisionZekai Wu, Shuqi Fan, Mengyin Liu, Yuhua Luo et al.CVPR 2026 · 2 citations
- Local-Global Coupling Spiking Graph Transformer for Brain Disorders Diagnosis from Two PerspectivesGeng Zhang, Jiangrong Shen, Kaizhong Zheng, Liangjun Chen et al.NeurIPS 2025 · 1 citation
- EvReflection: Event-Driven Micro-Dynamics for Reflection RemovalJiaxiao Wang, Dachun Kai, Huyue Zhu, Quanquan Hu et al.ICML 2026
- Robust Selective Activation with Randomized Temporal K-Winner-Take-All in Spiking Neural Networks for Continual LearningJiangrong Shen, Liang Zhao, Qi Xu, Yuqi Yang et al.ICLR 2026
- Bio-Vision-Inspired Spiking Neural Networks for Object Detection with Event CamerasDongyang Ma, Zhengyu Ma, Yifan Huang, Chenlin Zhou et al.ICML 2026
Builds on11
- Learning to Detect Objects with a 1 Megapixel Event CameraEtienne Perot, Pierre de Tournemire, Davide Nitti, Jonathan Masci et al.NeurIPS 2020 · 381 citations
- Deep Directly-Trained Spiking Neural Networks for Object DetectionQiaoyi Su, Yuhong Chou, Yifan Hu, Jianing Li et al.ICCV 2023 · 143 citations
- AEGNN: Asynchronous Event-based Graph Neural NetworksSimon Schaefer, Daniel Gehrig, Davide ScaramuzzaCVPR 2022 · 135 citations
- GET: Group Event Transformer for Event-Based VisionYansong Peng, Yueyi Zhang, Zhiwei Xiong, Xiaoyan Sun et al.ICCV 2023 · 86 citations
- From Chaos Comes Order: Ordering Event Representations for Object Recognition and DetectionNikola Zubic, Daniel Gehrig, Mathias Gehrig, Davide ScaramuzzaICCV 2023 · 71 citations
Related papers
- Spiking Transformers for Event-based Single Object TrackingJiqing Zhang, Bo Dong, Haiwei Zhang, Jianchuan Ding et al.CVPR 2022 · 171 citations
- Efficient Event-Based Object Detection: A Hybrid Neural Network with Spatial and Temporal AttentionSoikat Hasan Ahmed, Jan Finkbeiner, Emre NeftciCVPR 2025
- CREST: An Efficient Conjointly-trained Spike-driven Framework for Event-based Object Detection Exploiting Spatiotemporal DynamicsRuixin Mao, Aoyu Shen, Lin Tang, Jun ZhouAAAI 2025
- EGSST: Event-based Graph Spatiotemporal Sensitive Transformer for Object DetectionSheng Wu, Hang Sheng, Hui Feng, Bo HuNeurIPS 2024 · 9 citations
- Event-based Video Reconstruction Using TransformerWenming Weng, Yueyi Zhang, Zhiwei XiongICCV 2021 · 139 citations
