Collaboration Helps Camera Overtake LiDAR in 3D Detection
Yue Hu, Yifan Lu, Runsheng Xu, Weidi Xie, Siheng Chen, Yanfeng Wang
摘要
Camera-only 3D detection provides an economical solution with a simple configuration for localizing objects in 3D space compared to LiDAR-based detection systems. However, a major challenge lies in precise depth estimation due to the lack of direct 3D measurements in the input. Many previous methods attempt to improve depth estimation through network designs, e.g., deformable layers and larger receptive fields. This work proposes an orthogonal direction, improving the camera-only 3D detection by introducing multi-agent collaborations. Our proposed collaborative camera-only 3D detection (CoCa3D) enables agents to share complementary information with each other through communication. Meanwhile, we optimize communication efficiency by selecting the most informative cues. The shared messages from multiple viewpoints disambiguate the single-agent estimated depth and complement the occluded and long-range regions in the single-agent view. We evaluate CoCa3D in one real-world dataset and two new simulation datasets. Results show that CoCa3D improves previous SOTA performances by 44.21% on DAIR-V2X, 30.60% on OPV2V+, 12.59% on CoPerception-UAVs+ for AP@70. Our preliminary results show a potential that with sufficient collaboration, the camera might overtake LiDAR in some practical scenarios. We released the dataset and code.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper37
- How2comm: Communication-Efficient and Collaboration-Pragmatic Multi-Agent PerceptionDingkang Yang, Kun Yang, Yuzheng Wang, Jing Liu 等NeurIPS 2023 · 被引用 160 次
- An Extensible Framework for Open Heterogeneous Collaborative PerceptionYifan Lu, Yue Hu, Yiqi Zhong, Dequan Wang 等ICLR 2024 · 被引用 116 次
- Spatio-Temporal Domain Awareness for Multi-Agent Collaborative PerceptionKun Yang, Dingkang Yang, Jingyu Zhang, Mingcheng Li 等ICCV 2023 · 被引用 99 次
- TUMTraf V2X Cooperative Perception DatasetWalter Zimmer, Gerhard Arya Wardana, Suren Sritharan, Xingcheng Zhou 等CVPR 2024 · 被引用 76 次
- Flow-Based Feature Fusion for Vehicle-Infrastructure Cooperative 3D Object DetectionHaibao Yu, Yingjuan Tang, Enze Xie, Jilei Mao 等NeurIPS 2023 · 被引用 76 次
它引用的顶会 Paper12
- Voxel R-CNN: Towards High Performance Voxel-based 3D Object DetectionJiajun Deng, Shaoshuai Shi, Peiwei Li, Wengang Zhou 等AAAI 2021 · 被引用 1,128 次
- Where2comm: Communication-Efficient Collaborative Perception via Spatial Confidence MapsYue Hu, Shaoheng Fang, Zixing Lei, Yiqi Zhong 等NeurIPS 2022 · 被引用 537 次
- DAIR-V2X: A Large-Scale Dataset for Vehicle-Infrastructure Cooperative 3D Object DetectionHaibao Yu, Yizhen Luo, Mao Shu, Yiyi Huo 等CVPR 2022 · 被引用 475 次
- CIA-SSD: Confident IoU-Aware Single-Stage Object Detector From Point CloudWu Zheng, Weiliang Tang, Sijin Chen, Li Jiang 等AAAI 2021 · 被引用 335 次
- Cross-view Transformers for real-time Map-view Semantic SegmentationBrady Zhou, Philipp KrähenbühlCVPR 2022 · 被引用 279 次
相关 Paper
- Collaborative Semantic Occupancy Prediction with Hybrid Feature Fusion in Connected Automated VehiclesRui Song, Chenwei Liang, Hu Cao, Zhiran Yan 等CVPR 2024
- SparseAlign: a Fully Sparse Framework for Cooperative Object DetectionYunshuang Yuan, Yan Xia, Daniel Cremers, Monika SesterCVPR 2025
- Learning to Detect Objects from Multi-Agent LiDAR Scans without Manual LabelsQiming Xia, Wenkai Lin, Haoen Xiang, Xun Huang 等CVPR 2025
- Drones Help Drones: A Collaborative Framework for Multi-Drone Object Trajectory Prediction and BeyondZhechao Wang, Peirui Cheng, Minxing Chen, Pengju Tian 等NeurIPS 2024 · 被引用 34 次
- Unsupervised Multi-agent and Single-agent Perception from Cooperative ViewsHaochen Yang, Baolu Li, Lei Li, Delin Ren 等CVPR 2026 · 被引用 2 次
