Event2Vec: Processing neuromorphic events directly by representations in vector space
Wei Fang, Priyadarshini Panda
Abstract
Neuromorphic event cameras possess superior temporal resolution, power efficiency, and dynamic range compared to traditional cameras. However, their asynchronous and sparse data format poses a significant challenge for conventional deep learning methods. Most existing methods either densify events into frames, sacrificing their sparse asynchronous nature, or use irregular models that are less compatible with GPU acceleration. Inspired by word-to-vector models, we propose event2vec, a novel representation that allows Transformers to process events directly. We demonstrate the effectiveness of event2vec on the DVS Gesture, ASL-DVS, and DVS-Lip benchmarks, showing that event2vec is remarkably parameter-efficient, features high throughput and low latency, and achieves high accuracy even with an extremely low number of events or low spatial resolutions. These results show that sparse asynchronous event data can be directly integrated into high-throughput Transformer architectures, offering an efficient paradigm for real-time neuromorphic vision. The code is provided at https://github.com/ Intelligent-Computing-Lab-Panda/ event2vec .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d0e028be-6300-4b2b-98df-f9ae8ff3a6f8Cited by top-tier papers1
Ask how each one uses itBuilds on16
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Transformers are RNNs: Fast Autoregressive Transformers with Linear AttentionAngelos Katharopoulos, Apoorv Vyas, Nikolaos Pappas, François FleuretICML 2020 · 2,665 citations
- Train Short, Test Long: Attention with Linear Biases Enables Input Length ExtrapolationOfir Press, Noah A. Smith, Mike LewisICLR 2022 · 1,168 citations
- Point-BERT: Pre-training 3D Point Cloud Transformers with Masked Point ModelingXumin Yu, Lulu Tang, Yongming Rao, Tiejun Huang et al.CVPR 2022 · 684 citations
Related papers
- AEDNet: Asynchronous Event Denoising with Spatial-Temporal Correlation among Irregular DataHuachen Fang, Jinjian Wu, Leida Li, Junhui Hou et al.ACM MM 2022 · 26 citations
- Leveraging Asynchronous Spiking Neural Networks for Ultra Efficient Event-Based Visual ProcessingDingyi Zeng, Yuchen Wang, Honglin Cao, Wanlong Liu et al.AAAI 2025 · 2 citations
- GET: Group Event Transformer for Event-Based VisionYansong Peng, Yueyi Zhang, Zhiwei Xiong, Xiaoyan Sun et al.ICCV 2023 · 86 citations
- Graph-Based Object Classification for Neuromorphic Vision SensingYin Bi, Aaron Chadha, Alhabib Abbas, Eirina Bourtsoulatze et al.ICCV 2019 · 195 citations
- Stereo Depth from Events Cameras: Concentrate and Focus on the FutureYeongwoo Nam, S. Mohammad Mostafavi I., Kuk-Jin Yoon, Jonghyun ChoiCVPR 2022 · 56 citations
