Bring Event into RGB and LiDAR: Hierarchical Visual-Motion Fusion for Scene Flow
Hanyu Zhou, Yi Chang, Zhiwei Shi
摘要
Single RGB or LiDAR is the mainstream sensor for the challenging scene flow, which relies heavily on visual features to match motion features. Compared with single modality, existing methods adopt a fusion strategy to directly fuse the cross-modal complementary knowledge in motion space. However, these direct fusion methods may suffer the modality gap due to the visual intrinsic heterogeneous nature between RGB and LiDAR, thus deteriorating motion features. We dis-cover that event has the homogeneous nature with RGB and LiDAR in both visual and motion spaces. In this work, we bring the event as a bridge between RGB and LiDAR, and propose a novel hierarchical visual-motion fusion frame-work for scene flow, which explores a homogeneous space to fuse the cross-modal complementary knowledge for physical interpretation. In visual fusion, we discover that event has a complementarity (relative v.s. absolute) in luminance space with RGB for high dynamic imaging, and has a complemen-tarity (local boundary v.s. global shape) in scene structure space with LiDAR for structure integrity. In motion fusion, we figure out that RGB, event and LiDAR are complementary (spatial-dense, temporal-dense v.s. spatiotemporal-sparse) to each other in correlation space, which motivates us to fuse their motion correlations for motion continuity. The proposed hierarchical fusion can explicitly fuse the multimodal knowledge to progressively improve scene flow from visual space to motion space. Extensive experiments have been performed to verify the superiority of the proposed method.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Injecting Frame-Event Complementary Fusion into Diffusion for Optical Flow in Challenging ScenesHaonan Wang, Hanyu Zhou, Haoyue Liu, Luxin YanNeurIPS 2025 · 被引用 4 次
- DuCos: Duality Constrained Depth Super-Resolution via Foundation ModelZhiqiang Yan, Zhengxue Wang, Haoye Dong, Jun Li 等ICCV 2025 · 被引用 3 次
- Emotive: Event-Guided Trajectory Modeling for 3D Motion EstimationZengyu Wan, Wei Zhai, Yang Cao, Zhengjun ZhaICCV 2025 · 被引用 1 次
- PEOD: A Pixel-Aligned Event-RGB Benchmark for Object Detection Under Challenging ConditionsLuoping Cui, Hanqing Liu, Mingjie Liu, Endian Lin 等AAAI 2026 · 被引用 1 次
- Event-Aided Dense and Continuous Point Tracking: Everywhere and AnytimeZhexiong Wan, Jianqin Luo, Yuchao Dai, Gim Hee LeeICCV 2025 · 被引用 1 次
它引用的顶会 Paper17
- Learning to Estimate Hidden Motions with Global Motion AggregationShihao Jiang, Dylan Campbell, Yao Lu, Hongdong Li 等ICCV 2021 · 被引用 402 次
- RGB-D Saliency Detection via Cascaded Mutual Information MinimizationJing Zhang, Deng-Ping Fan, Yuchao Dai, Xin Yu 等ICCV 2021 · 被引用 122 次
- SLIM: Self-Supervised LiDAR Scene Flow and Motion SegmentationStefan Andreas Baur, David Josef Emmerichs, Frank Moosmann, Peter Pinggera 等ICCV 2021 · 被引用 110 次
- ObjectFusion: Multi-modal 3D Object Detection with Object-Centric FusionQi Cai, Yingwei Pan, Ting Yao, Chong-Wah Ngo 等ICCV 2023 · 被引用 71 次
- CamLiFlow: Bidirectional Camera-LiDAR Fusion for Joint Optical Flow and Scene Flow EstimationHaisong Liu, Tao Lu, Yihui Xu, Jia Liu 等CVPR 2022 · 被引用 64 次
相关 Paper
- Bridge Frame and Event: Common Spatiotemporal Fusion for High-Dynamic Scene Optical FlowHanyu Zhou, Haonan Wang, Haoyue Liu, Yuxing Duan 等CVPR 2025
- x^2-Fusion: Cross-Modality and Cross-Dimension Flow Estimation in Event Edge SpaceRuishan Guo, Ciyu Ruan, Haoyang Wang, Zihang Gong 等CVPR 2026
- RPEFlow: Multimodal Fusion of RGB-PointCloud-Event for Joint Optical Flow and Scene Flow EstimationZhexiong Wan, Yuxin Mao, Jing Zhang, Yuchao DaiICCV 2023 · 被引用 35 次
- RaLiFlow: Scene Flow Estimation with 4D Radar and LiDAR Point CloudsJingyun Fu, Zhiyu Xiang, Na ZhaoAAAI 2026
- ARES: Unifying Asymmetric RGB-Event Stereo for Probabilistic Scene Flow EstimationJie Long Lee, Gim Hee LeeCVPR 2026
