RoCo: Robust Cooperative Perception By Iterative Object Matching and Pose Adjustment
Zhe Huang, Shuo Wang, Yongcai Wang, Wanting Li, Deying Li, Lei Wang
摘要
Collaborative autonomous driving with multiple vehicles usually requires the data fusion from multiple modalities. To ensure effective fusion, the data from each individual modality shall maintain a reasonably high quality. However, in collaborative perception, the quality of object detection based on a modality is highly sensitive to the relative pose errors among the agents. It leads to feature misalignment and significantly reduces collaborative performance. To address this issue, we propose RoCo, a novel unsupervised framework to conduct iterative object matching and agent pose adjustment. To the best of our knowledge, our work is the first to model the pose correction problem in collaborative perception as an object matching task, which reliably associates common objects detected by different agents. On top of this, we propose a graph optimization process to adjust the agent poses by minimizing the alignment errors of the associated objects, and the object matching is re-done based on the adjusted agent poses. This process is carried out iteratively until convergence. Experimental study on both simulated and real-world datasets demonstrates that the proposed framework RoCo consistently outperforms existing relevant methods in terms of the collaborative object detection performance, and exhibits highly desired robustness when the pose information of agents is with high-level noise. Ablation studies are also provided to show the impact of its key parameters and components. The code is released at https://github.com/HuangZhe885/RoCo.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- CATNet: Collaborative Alignment and Transformation Network for Cooperative PerceptionGong Chen, Chaokun Zhang, Tao Tang, Pengcheng Lv 等CVPR 2026 · 被引用 1 次
- Point-Cache: Test-time Dynamic and Hierarchical Cache for Robust and Generalizable Point Cloud AnalysisHongyu Sun, Qiuhong Ke, Ming Cheng, Yongcai Wang 等CVPR 2025
- MambaVO: Deep Visual Odometry Based on Sequential Matching Refinement and Training SmoothingShuo Wang, Wanting Li, Yongcai Wang, Zhaoxin Fan 等CVPR 2025
它引用的顶会 Paper10
- Where2comm: Communication-Efficient Collaborative Perception via Spatial Confidence MapsYue Hu, Shaoheng Fang, Zixing Lei, Yiqi Zhong 等NeurIPS 2022 · 被引用 537 次
- DAIR-V2X: A Large-Scale Dataset for Vehicle-Infrastructure Cooperative 3D Object DetectionHaibao Yu, Yizhen Luo, Mao Shu, Yiyi Huo 等CVPR 2022 · 被引用 475 次
- An Extensible Framework for Open Heterogeneous Collaborative PerceptionYifan Lu, Yue Hu, Yiqi Zhong, Dequan Wang 等ICLR 2024 · 被引用 116 次
- Asynchrony-Robust Collaborative Perception via Bird's Eye View FlowSizhe Wei, Yuxi Wei, Yue Hu, Yifan Lu 等NeurIPS 2023 · 被引用 102 次
- What2comm: Towards Communication-efficient Collaborative Perception via Feature DecouplingKun Yang, Dingkang Yang, Jingyu Zhang, Hanqi Wang 等ACM MM 2023 · 被引用 58 次
相关 Paper
- FeaCo: Reaching Robust Feature-Level Consensus in Noisy Pose ConditionsJiaming Gu, Jingyu Zhang, Muyang Zhang, Weiliang Meng 等ACM MM 2023 · 被引用 21 次
- mmCooper: A Multi-Agent Multi-Stage Communication-Efficient and Collaboration-Robust Cooperative Perception FrameworkBingyi Liu, Jian Teng, Hongfei Xue, Enshu Wang 等ICCV 2025 · 被引用 14 次
- RoCo-Sim: Enhancing Roadside Collaborative Perception through Foreground SimulationYuwen Du, Anning Hu, Zichen Chao, Yifan Lu 等ICCV 2025 · 被引用 2 次
- Multi-Agent Collaborative Perception via Motion-Aware Robust Communication NetworkShixin Hong, Yu Liu, Zhi Li, Shaohui Li 等CVPR 2024
- GT-Space: Enhancing Heterogeneous Collaborative Perception with Ground Truth Feature SpaceWentao Wang, Haoran Xu, Guang TanICLR 2026 · 被引用 2 次
