SparseInteraction: Sparse Semantic Guidance for Radar and Camera 3D Object Detection
Shaoqing Xu, Shengyin Jiang, Fang Li, Li Liu, Ziying Song, Bo Yang, Zhixin Yang
摘要
Multi-modal fusion techniques, such as radar and images, enable a complementary and cost-effective perception of the surrounding environment regardless of lighting and weather conditions. However, existing fusion methods for surround-view images and radar are challenged by the inherent noise and positional ambiguity of radar, which leads to significant performance losses. To address this limitation effectively, our paper presents a robust, end-to-end fusion framework dubbed SparseInteraction. First, we introduce the Noisy Radar Filter (NRF) module to extract foreground features by creatively using queried semantic features from the image to filter out noisy radar features. Furthermore, we implement the Sparse Cross-Attention Encoder (SCAE) to effectively blend foreground radar features and image features to address positional ambiguity issues at a sparse level. Ultimately, to facilitate model convergence and performance, the foreground prior queries containing position information of the foreground radar are concatenated with predefined queries and fed into the subsequent transformer-based decoder. The experimental results demonstrate that the proposed fusion strategies markedly enhance detection performance and achieve new state-of-the-art results on the nuScenes benchmark. Source code is available at https://github.com/GG-Bonds/SparseInteraction.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- CRAFT: Camera-Radar 3D Object Detection with Spatio-Contextual Fusion TransformerYoungseok Kim, Sanmin Kim, Jun Won Choi, Dongsuk KumAAAI 2023 · 被引用 145 次
- RCTrans: Radar-Camera Transformer via Radar Densifier and Sequential Decoder for 3D Object DetectionYiheng Li, Yang Yang, Zhen LeiAAAI 2025 · 被引用 4 次
- CRN: Camera Radar Net for Accurate, Robust, Efficient 3D PerceptionYoungseok Kim, Juyeb Shin, Sanmin Kim, In-Jae Lee 等ICCV 2023 · 被引用 134 次
- SparseFusion: Fusing Multi-Modal Sparse Representations for Multi-Sensor 3D Object DetectionYichen Xie, Chenfeng Xu, Marie-Julie Rakotosaona, Patrick Rim 等ICCV 2023 · 被引用 134 次
- RC-AutoCalib: An End-to-End Radar-Camera Automatic Calibration NetworkVan-Tin Luu, Yon-Lin Cai, Vu-Hoang Tran, Wei-Chen Chiu 等CVPR 2025
