Towards Robust Event-Based Depth Estimation: Bridging Synthetic and Real Domains with Motion Adaptation
Yuzhe Ji, Haotian Wang, Yijie Chen, Xiang Cheng, Liuqing Yang, Xinhu Zheng
摘要
Event cameras offer microsecond latency and high dynamic range, making them particularly suitable for safety-critical 3D perception in autonomous driving scenarios with challenging lighting conditions. Yet existing methods often struggle to generalize to out-of-domain environments due to the limited availability of diverse training data. While synthetic data offers an easily accessible alternative, it introduces a significant sim-to-real gap, particularly in motion patterns. We tackle this challenge by introducing Motion-Adaptation Mamba (MA-Mamba), a dual-track framework that advances both architecture and data augmentation. At the architectural level, we introduce a lightweight Spatial-Temporal Association module that captures motion-induced appearance variations at arbitrary scales, and an Adaptive Memory Balancing module, built on the Mamba state-space framework, that adaptively filters memory updates to maintain stable scene context under diverse dynamics. At the data level, we design event-oriented augmentations that simulate varied motion patterns and apply priority-based masked sequence modeling to strengthen long-range spatio-temporal reasoning. Trained solely on synthetic data, MA-Mamba delivers substantial zero-shot gains on multiple real-world benchmarks, demonstrating strong robustness and generalizability.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper17
- Searching for MobileNetV3Andrew Howard, Ruoming Pang, Hartwig Adam, Quoc V. Le 等ICCV 2019 · 被引用 9,163 次
- BEiT: BERT Pre-Training of Image TransformersHangbo Bao, Li Dong, Songhao Piao, Furu WeiICLR 2022 · 被引用 3,632 次
- Efficiently Modeling Long Sequences with Structured State SpacesAlbert Gu, Karan Goel, Christopher RéICLR 2022 · 被引用 3,482 次
- Depth Anything: Unleashing the Power of Large-Scale Unlabeled DataLihe Yang, Bingyi Kang, Zilong Huang, Xiaogang Xu 等CVPR 2024 · 被引用 847 次
- On the Parameterization and Initialization of Diagonal State Space ModelsAlbert Gu, Karan Goel, Ankit Gupta, Christopher RéNeurIPS 2022 · 被引用 690 次
相关 Paper
- EventDrive: Event Cameras for Vision-Language Driving IntelligenceDongyue Lu, Rong Li, Ao Liang, Lingdong Kong 等CVPR 2026 · 被引用 2 次
- MambaSeg: Harnessing Mamba for Accurate and Efficient Image-Event Semantic SegmentationFuqiang Gu, Yuanke Li, Xianlei Long, Kangping Ji 等AAAI 2026 · 被引用 1 次
- An Event-tailored State-Space Based Model for Pedestrian DetectionLiuyi Li, Feng Shi, Jian Wang, Jinjing Zhu 等ACM MM 2025
- E-MaT: Event-oriented Mamba for Egocentric Point TrackingHan Han, Wei Zhai, Baocai Yin, Yang Cao 等AAAI 2026
- EA3D: Event-Augmented 3D Diffusion for Generalizable Novel View SynthesisWangbo Yu, Chaoran Feng, Jianing Li, Aofan Zhang 等ICLR 2026
