Depth Any Event Stream: Enhancing Event-based Monocular Depth Estimation via Dense-to-Sparse Distillation
Jinjing Zhu, Tianbo Pan, Zidong Cao, Yexin Liu, James T. Kwok, Hui Xiong
摘要
With the superior sensitivity of event cameras to high-speed motion and extreme lighting conditions, event-based monocular depth estimation has gained popularity to predict structural information about surrounding scenes in challenging environments. However, the scarcity of labeled event data constrains prior supervised learning methods. Unleashing the promising potential of the existing RGB-based depth foundation model, DAM, we propose Depth Any Event stream (EventDAM) to achieve high-performance event based monocular depth estimation in an annotation-free manner. EventDAM effectively combines paired dense RGB images with sparse event data by incorporating three key cross-modality components: Sparsity-aware Feature Mixture (SFM), Sparsity-aware Feature Distillation (SFD), and Sparsity-invariant Consistency Module (SCM). With the proposed sparsity metric, SFM mixes features from RGB images and event data to generate auxiliary depth predictions, while SFD facilitates adaptive feature distillation. Furthermore, SCM ensures output consistency across varying sparsity levels in event data, thereby endowing EventDAM with zero shot capabilities across diverse scenes. Extensive experiments across a variety of benchmark datasets, compared to approaches using diverse input modalities, robustly substantiate the generalization and zero-shot capabilities of EventDAM.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Event6D: Event-based Novel Object 6D Pose TrackingJae-Young Kang, Hoonhee Cho, Taeyeop Lee, Minjun Kang 等CVPR 2026 · 被引用 4 次
- Scaling Dense Event-Stream Pretraining from Visual Foundation ModelsZhiwen Chen, Junhui Hou, Zhiyu Zhu, Jinjian Wu 等CVPR 2026 · 被引用 2 次
- SkyEvents: A Large-Scale Event-enhanced UAV Dataset for Robust 3D Scene ReconstructionWenzong Ma, Zhuoxiao Li, Jinjing Zhu, Tongyan Hua 等ICLR 2026
- EvDiff3D: Event-Aware Diffusion Repair for High-Fidelity Event-Based 3D ReconstructionKanghao Chen, Zixin Zhang, Hangyu Li, Lin Wang 等AAAI 2026
- EventHub: Data Factory for Generalizable Event-Based Stereo Networks without Active SensorsLuca Bartolomei, Fabio Tosi, Matteo Poggi, Stefano Mattoccia 等CVPR 2026
它引用的顶会 Paper21
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao 等ICCV 2023 · 被引用 13,211 次
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou 等ICCV 2021 · 被引用 8,921 次
- Depth Anything V2Lihe Yang, Bingyi Kang, Zilong Huang, Zhen Zhao 等NeurIPS 2024 · 被引用 2,305 次
相关 Paper
- Depth AnyEvent: A Cross-Modal Distillation Paradigm for Event-Based Monocular Depth EstimationLuca Bartolomei, Enrico Mannocci, Fabio Tosi, Matteo Poggi 等ICCV 2025 · 被引用 13 次
- Distil-E2D: Distilling Image-to-Depth Priors for Event-Based Monocular Depth EstimationJie Long Lee, Gim Hee LeeNeurIPS 2025 · 被引用 3 次
- Enhanced Event-Based Dense Stereo via Cross-Sensor Knowledge DistillationHaihao Zhang, Yunjian Zhang, Jianing Li, Lin Zhu 等ICCV 2025 · 被引用 1 次
- Unsupervised 3d Motion Estimation Using Event CameraHan Han, Wei Zhai, Tiesong Zhao, Bin Li 等CVPR 2026
- Event-Image Fusion Stereo Using Cross-Modality Feature PropagationHoonhee Cho, Kuk-Jin YoonAAAI 2022 · 被引用 34 次
