Bring Event into RGB and LiDAR: Hierarchical Visual-Motion Fusion for Scene Flow
Hanyu Zhou, Yi Chang, Zhiwei Shi
Abstract
Single RGB or LiDAR is the mainstream sensor for the challenging scene flow, which relies heavily on visual features to match motion features. Compared with single modality, existing methods adopt a fusion strategy to directly fuse the cross-modal complementary knowledge in motion space. However, these direct fusion methods may suffer the modality gap due to the visual intrinsic heterogeneous nature between RGB and LiDAR, thus deteriorating motion features. We dis-cover that event has the homogeneous nature with RGB and LiDAR in both visual and motion spaces. In this work, we bring the event as a bridge between RGB and LiDAR, and propose a novel hierarchical visual-motion fusion frame-work for scene flow, which explores a homogeneous space to fuse the cross-modal complementary knowledge for physical interpretation. In visual fusion, we discover that event has a complementarity (relative v.s. absolute) in luminance space with RGB for high dynamic imaging, and has a complemen-tarity (local boundary v.s. global shape) in scene structure space with LiDAR for structure integrity. In motion fusion, we figure out that RGB, event and LiDAR are complementary (spatial-dense, temporal-dense v.s. spatiotemporal-sparse) to each other in correlation space, which motivates us to fuse their motion correlations for motion continuity. The proposed hierarchical fusion can explicitly fuse the multimodal knowledge to progressively improve scene flow from visual space to motion space. Extensive experiments have been performed to verify the superiority of the proposed method.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 98ab82ae-89fa-45ed-a1f0-3a0f03cb86b4Cited by top-tier papers8
- Injecting Frame-Event Complementary Fusion into Diffusion for Optical Flow in Challenging ScenesHaonan Wang, Hanyu Zhou, Haoyue Liu, Luxin YanNeurIPS 2025 · 4 citations
- DuCos: Duality Constrained Depth Super-Resolution via Foundation ModelZhiqiang Yan, Zhengxue Wang, Haoye Dong, Jun Li et al.ICCV 2025 · 3 citations
- Emotive: Event-Guided Trajectory Modeling for 3D Motion EstimationZengyu Wan, Wei Zhai, Yang Cao, Zhengjun ZhaICCV 2025 · 1 citation
- PEOD: A Pixel-Aligned Event-RGB Benchmark for Object Detection Under Challenging ConditionsLuoping Cui, Hanqing Liu, Mingjie Liu, Endian Lin et al.AAAI 2026 · 1 citation
- Event-Aided Dense and Continuous Point Tracking: Everywhere and AnytimeZhexiong Wan, Jianqin Luo, Yuchao Dai, Gim Hee LeeICCV 2025 · 1 citation
Builds on17
- Learning to Estimate Hidden Motions with Global Motion AggregationShihao Jiang, Dylan Campbell, Yao Lu, Hongdong Li et al.ICCV 2021 · 402 citations
- RGB-D Saliency Detection via Cascaded Mutual Information MinimizationJing Zhang, Deng-Ping Fan, Yuchao Dai, Xin Yu et al.ICCV 2021 · 122 citations
- SLIM: Self-Supervised LiDAR Scene Flow and Motion SegmentationStefan Andreas Baur, David Josef Emmerichs, Frank Moosmann, Peter Pinggera et al.ICCV 2021 · 110 citations
- ObjectFusion: Multi-modal 3D Object Detection with Object-Centric FusionQi Cai, Yingwei Pan, Ting Yao, Chong-Wah Ngo et al.ICCV 2023 · 71 citations
- CamLiFlow: Bidirectional Camera-LiDAR Fusion for Joint Optical Flow and Scene Flow EstimationHaisong Liu, Tao Lu, Yihui Xu, Jia Liu et al.CVPR 2022 · 64 citations
Related papers
- Bridge Frame and Event: Common Spatiotemporal Fusion for High-Dynamic Scene Optical FlowHanyu Zhou, Haonan Wang, Haoyue Liu, Yuxing Duan et al.CVPR 2025
- x^2-Fusion: Cross-Modality and Cross-Dimension Flow Estimation in Event Edge SpaceRuishan Guo, Ciyu Ruan, Haoyang Wang, Zihang Gong et al.CVPR 2026
- RPEFlow: Multimodal Fusion of RGB-PointCloud-Event for Joint Optical Flow and Scene Flow EstimationZhexiong Wan, Yuxin Mao, Jing Zhang, Yuchao DaiICCV 2023 · 35 citations
- RaLiFlow: Scene Flow Estimation with 4D Radar and LiDAR Point CloudsJingyun Fu, Zhiyu Xiang, Na ZhaoAAAI 2026
- ARES: Unifying Asymmetric RGB-Event Stereo for Probabilistic Scene Flow EstimationJie Long Lee, Gim Hee LeeCVPR 2026
