RPGFusion: 4D Radar Prior-Guided Multi-Modal Fusion for 3D Detection
Xin Qiu, Wenjie Liu
Abstract
Accurate 3D object detection in autonomous driving relies on effectively combining complementary information from multiple sensors. 4D millimeter-wave radar provides sparse yet physically reliable measurements, whose potential for enhancing sensor fusion has not been fully utilized. In this work, we propose Radar Prior Guided Fusion (RPGFusion), a practical 4D radar–camera fusion framework. We first generate radar prior maps that encode spatial confidence and depth cues. These priors guide image feature sampling while preventing the uneven BEV feature distribution (near-dense, far-sparse) caused by Lift-Splat-Shoot view transformation. To address the sparsity and noise inherent in point clouds, we adopt a hybrid robust encoding and sparse-to-dense feature propagation. We further introduce spatial alignment and semantic fusion modules to reconcile geometric and semantic differences between modalities, yielding more consistent and complementary BEV representations. Extensive experiments on the public View-of-Delft and TJ4DRadSet show that RPGFusion outperforms prior radar–camera fusion methods, achieving SOTA performance. Our work not only uses 4D radar signals to guide image BEV queries, but also enables robust radar feature encoding and densification for 3D perception, demonstrating the strong potential of 4D radar.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on17
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- BEVDepth: Acquisition of Reliable Depth for Multi-View 3D Object DetectionYinhao Li, Zheng Ge, Guanyi Yu, Jinrong Yang et al.AAAI 2023 · 954 citations
- CRAFT: Camera-Radar 3D Object Detection with Spatio-Contextual Fusion TransformerYoungseok Kim, Sanmin Kim, Jun Won Choi, Dongsuk KumAAAI 2023 · 145 citations
- CRN: Camera Radar Net for Accurate, Robust, Efficient 3D PerceptionYoungseok Kim, Juyeb Shin, Sanmin Kim, In-Jae Lee et al.ICCV 2023 · 134 citations
- 3D Point Cloud Generation with Millimeter-Wave RadarKun Qian, Zhaoyuan He, Xinyu ZhangUbiComp 2021 · 111 citations
Related papers
- CVFusion: Cross-View Fusion of 4D Radar and Camera for 3D Object DetectionHanzhi Zhong, Zhiyu Xiang, Ruoyu Xu, Jingyun Fu et al.ICCV 2025 · 5 citations
- R4Det: 4D Radar-Camera Fusion for High-Performance 3D Object DetectionZhongyu Xia, Yousen Tang, Yongtao Wang, Zhifeng Wang et al.CVPR 2026 · 4 citations
- RaGS: Unleashing 3D Gaussian Splatting from 4D Radar and Monocular Cue for 3D Object DetectionXiaokai Bai, Chenxu Zhou, Lianqing Zheng, Jianan Liu et al.CVPR 2026
- HGSFusion: Radar-Camera Fusion with Hybrid Generation and Synchronization for 3D Object DetectionZijian Gu, Jianwei Ma, Yan Huang, Honghao Wei et al.AAAI 2025 · 26 citations
- Doppler-Aware LiDAR-RADAR Fusion for Weather-Robust 3D DetectionYujeong Chae, Heejun Park, Hyeonseong Kim, Kuk-Jin YoonICCV 2025 · 6 citations
