E2PNet: Event to Point Cloud Registration with Spatio-Temporal Representation Learning
Xiuhong Lin, Changjie Qiu, Zhipeng Cai, Siqi Shen, Yu Zang, Weiquan Liu, Xuesheng Bian, Matthias Müller, Cheng Wang
Abstract
Event cameras have emerged as a promising vision sensor in recent years due to their unparalleled temporal resolution and dynamic range. While registration of 2D RGB images to 3D point clouds is a long-standing problem in computer vision, no prior work studies 2D-3D registration for event cameras. To this end, we propose E2PNet, the first learning-based method for event-to-point cloud registration. The core of E2PNet is a novel feature representation network called Event-Points-to-Tensor (EP2T), which encodes event data into a 2D grid-shaped feature tensor. This grid-shaped feature enables matured RGB-based frameworks to be easily used for event-to-point cloud registration, without changing hyper-parameters and the training procedure. EP2T treats the event input as spatio-temporal point clouds. Unlike standard 3D learning architectures that treat all dimensions of point clouds equally, the novel sampling and information aggregation modules in EP2T are designed to handle the inhomogeneity of the spatial and temporal dimensions. Experiments on the MVSEC and VECtor datasets demonstrate the superiority of E2PNet over hand-crafted and other learning-based methods. Compared to RGB-based registration, E2PNet is more robust to extreme illumination or fast motion due to the use of event data. Beyond 2D-3D registration, we also show the potential of EP2T for other vision tasks such as flow estimation, event-to-image reconstruction and object recognition. The source code can be found at: https://github.com/Xmu-qcj/E2PNet.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers5
- Event-3DGS: Event-based 3D Reconstruction Using 3D Gaussian SplattingHaiqian Han, Jianing Li, Henglu Wei, Xiangyang JiNeurIPS 2024 · 39 citations
- RELI11D: A Comprehensive Multimodal Human Motion Dataset and MethodMing Yan, Yan Zhang, Shuqiang Cai, Shuqi Fan et al.CVPR 2024 · 5 citations
- EventFlash: Towards Efficient MLLMs for Event-Based VisionShaoyu Liu, Jianing Li, Guanghui Zhao, Yunjian Zhang et al.ICLR 2026 · 5 citations
- Event2Vec: Processing neuromorphic events directly by representations in vector spaceWei Fang, Priyadarshini PandaICML 2026 · 4 citations
- OmniEvent: Unified Event Representation LearningWeiqi Yan, Chenlu Lin, Youbiao Wang, Zhipeng Cai et al.AAAI 2026
Builds on12
- End-to-End Learning of Representations for Asynchronous Event-Based DataDaniel Gehrig, Antonio Loquercio, Konstantinos G. Derpanis, Davide ScaramuzzaICCV 2019 · 427 citations
- AEGNN: Asynchronous Event-based Graph Neural NetworksSimon Schaefer, Daniel Gehrig, Davide ScaramuzzaCVPR 2022 · 135 citations
- Graph-based Asynchronous Event Processing for Rapid Object RecognitionYijin Li, Han Zhou, Bangbang Yang, Ye Zhang et al.ICCV 2021 · 131 citations
- Learning an Event Sequence Embedding for Dense Event-Based Deep StereoStepan Tulyakov, François Fleuret, Martin Kiefel, Peter V. Gehler et al.ICCV 2019 · 122 citations
- LCD: Learned Cross-Domain Descriptors for 2D-3D MatchingQuang-Hieu Pham, Mikaela Angelina Uy, Binh-Son Hua, Duc Thanh Nguyen et al.AAAI 2020 · 94 citations
Related papers
- Scalable Event Cloud Network for Event-based ClassificationHongwei Ren, Fei Ma, Xiaopeng LIN, Yuetong Fang et al.ICML 2026 · 5 citations
- Dual Memory Aggregation Network for Event-Based Object Detection with Learnable RepresentationDongsheng Wang, Xu Jia, Yang Zhang, Xinyu Zhang et al.AAAI 2023 · 22 citations
- Dual Transfer Learning for Event-based End-task Prediction via Pluggable Event to Image TranslationLin Wang, Yujeong Chae, Kuk-Jin YoonICCV 2021 · 46 citations
- A Simple and Effective Point-Based Network for Event Camera 6-DOFs Pose RelocalizationHongwei Ren, Jiadong Zhu, Yue Zhou, Haotian Fu et al.CVPR 2024 · 14 citations
- Deep Event Stereo Leveraged by Event-to-Image TranslationSoikat Hasan Ahmed, Hae Woong Jang, S. M. Nadim Uddin, Yong Ju JungAAAI 2021 · 41 citations
