Weakly Supervised Cross-Modal Learning for 4D Radar Scene Flow Estimation
Jingyun Fu, Zhiyu Xiang, Na Zhao
Abstract
Due to the difficulty of obtaining ground-truth data for 4D radar scene flow estimation, previous methods typically rely on either self-supervised losses or cross-modal supervision using 3D LiDAR data, 2D images, and odometry. However, self-supervised approaches often yield suboptimal results due to radar's inherently low-fidelity measurements, while existing cross-modal supervised methods introduce complex multi-task architecture and require costly LiDAR sensors to generate pseudo radar scene flow labels from pretrained 3D tracking models. To overcome these limitations, we propose a task-specific iterative framework for weakly supervised radar scene flow learning, using only images and odometry for auxiliary supervision during training. Specially, we establish two novel instance-aware self-supervised losses by exploiting off-the-shelf 2D tracking and segmentation algorithms to obtain tracked instance masks, which are back-projected into 3D space to provide instance-level semantic guidance; for static regions, we integrate vehicle odometry with radar's intrinsic motion cues to construct a rigid static loss. Extensive experiments on the real-world View-of-Delft (VoD) dataset demonstrate that our method not only surpasses state-of-the-art cross-modal supervised approaches that rely on 3D multi-object tracking on dense LiDAR point clouds but also outperforms existing fully supervised scene flow estimation methods. The code is open-sourced at https://github.com/FuJingyun/IterFlow.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2d461dae-9268-45df-adab-4f21214fe577Builds on20
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao et al.ICCV 2023 · 13,211 citations
- Neural Scene Flow PriorXueqian Li, Jhony Kaesemodel Pontes, Simon LuceyNeurIPS 2021 · 136 citations
- CamLiFlow: Bidirectional Camera-LiDAR Fusion for Joint Optical Flow and Scene Flow EstimationHaisong Liu, Tao Lu, Yihui Xu, Jia Liu et al.CVPR 2022 · 64 citations
- Fast Neural Scene FlowXueqian Li, Jianqiao Zheng, Francesco Ferroni, Jhony Kaesemodel Pontes et al.ICCV 2023 · 41 citations
- RPEFlow: Multimodal Fusion of RGB-PointCloud-Event for Joint Optical Flow and Scene Flow EstimationZhexiong Wan, Yuxin Mao, Jing Zhang, Yuchao DaiICCV 2023 · 35 citations
Related papers
- Hidden Gems: 4D Radar Scene Flow Learning Using Cross-Modal SupervisionFangqiang Ding, Andras Palffy, Dariu M. Gavrila, Chris Xiaoxuan LuCVPR 2023
- RaLiFlow: Scene Flow Estimation with 4D Radar and LiDAR Point CloudsJingyun Fu, Zhiyu Xiang, Na ZhaoAAAI 2026
- FGRFlow: Learning Fine-Grained Rigidity Scene Flow from 4D Radar Point CloudMingliang Zhai, Yiheng Wang, Haidong Hu, Chi-Man Pun et al.ACM MM 2025
- TARS: Traffic-Aware Radar Scene Flow EstimationJialong Wu, Marco Braun, Dominic Spata, Matthias RottmannICCV 2025
- CVFusion: Cross-View Fusion of 4D Radar and Camera for 3D Object DetectionHanzhi Zhong, Zhiyu Xiang, Ruoyu Xu, Jingyun Fu et al.ICCV 2025 · 5 citations
