A Simple and Effective Point-Based Network for Event Camera 6-DOFs Pose Relocalization
Hongwei Ren, Jiadong Zhu, Yue Zhou, Haotian Fu, Yulong Huang, Bojun Cheng
摘要
Event cameras exhibit remarkable attributes such as high dynamic range, asynchronicity, and low latency, making them highly suitable for vision tasks that involve highspeed motion in challenging lighting conditions. These cameras implicitly capture movement and depth information in events, making them appealing sensors for Camera Pose Relocalization (CPR) tasks. Nevertheless, existing CPR networks based on events neglect the pivotal finegrained temporal information in events, resulting in unsatisfactory performance. Moreover, the energy-efficient features are further compromised by the use of excessively complex models, hindering efficient deployment on edge devices. In this paper, we introduce PEPNet, a simple and effective point-based network designed to regress six degrees of freedom (6-DOFs) event camera poses. We rethink the relationship between the event camera and CPR tasks, leveraging the raw Point Cloud directly as network input to harness the high-temporal resolution and inherent sparsity of events. PEPNet is adept at abstracting the spatial and implicit temporal features through hierarchical structure and explicit temporal features by Attentive Bidirectional Long Short-Term Memory (A-Bi-LSTM). Byemploying a carefully crafted lightweight design, PEPNet delivers state-of-the-art (SOTA) performance on both indoor and outdoor datasets with meager computational resources. Specifically, PEPNet attains a significant 38% and 33% performance improvement on the random split IJRR and M3ED datasets, respectively. Moreover, the lightweight design version PEPNet<inf xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">tiny</inf> accomplishes results comparable to the SOTA while employing a mere 0.5% of the parameters.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Towards Robust Event-Based Depth Estimation: Bridging Synthetic and Real Domains with Motion AdaptationYuzhe Ji, Haotian Wang, Yijie Chen, Xiang Cheng 等AAAI 2026
- Focal Plane Visual Feature Generation and Matching on a Pixel Processor ArrayHongyi Zhang, Laurie Bose, Jianing Chen, Piotr Dudek 等ICCV 2025
它引用的顶会 Paper6
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Rethinking Network Design and Local Geometry in Point Cloud: A Simple Residual MLP FrameworkXu Ma, Can Qin, Haoxuan You, Haoxi Ran 等ICLR 2022 · 被引用 841 次
- SpikePoint: An Efficient Point-based Spiking Neural Network for Event Cameras Action RecognitionHongwei Ren, Yue Zhou, Xiaopeng Lin, Yulong Huang 等ICLR 2024 · 被引用 39 次
- Point TransformerHengshuang Zhao, Li Jiang, Jiaya Jia, Philip H. S. Torr 等ICCV 2021 · 被引用 23 次
- TTPOINT: A Tensorized Point Cloud Network for Lightweight Action Recognition with Event CamerasHongwei Ren, Yue Zhou, Haotian Fu, Yulong Huang 等ACM MM 2023 · 被引用 14 次
相关 Paper
- Recurrent Vision Transformers for Object Detection with Event CamerasMathias Gehrig, Davide ScaramuzzaCVPR 2023
- E2PNet: Event to Point Cloud Registration with Spatio-Temporal Representation LearningXiuhong Lin, Changjie Qiu, Zhipeng Cai, Siqi Shen 等NeurIPS 2023 · 被引用 18 次
- Event-based Video Reconstruction Using TransformerWenming Weng, Yueyi Zhang, Zhiwei XiongICCV 2021 · 被引用 139 次
- DERD-Net: Learning Depth from Event-based Ray DensitiesDiego de Oliveira Hitzges, Suman Ghosh, Guillermo GallegoNeurIPS 2025 · 被引用 6 次
- Active Event-based Stereo VisionJianing Li, Yunjian Zhang, Haiqian Han, Xiangyang JiCVPR 2025
