Learning Distilled Collaboration Graph for Multi-Agent Perception
Yiming Li, Shunli Ren, Pengxiang Wu, Siheng Chen, Chen Feng, Wenjun Zhang
Abstract
To promote better performance-bandwidth trade-off for multi-agent perception, we propose a novel distilled collaboration graph (DiscoGraph) to model trainable, pose-aware, and adaptive collaboration among agents. Our key novelties lie in two aspects. First, we propose a teacher-student framework to train DiscoGraph via knowledge distillation. The teacher model employs an early collaboration with holistic-view inputs; the student model is based on intermediate collaboration with single-view inputs. Our framework trains DiscoGraph by constraining post-collaboration feature maps in the student model to match the correspondences in the teacher model. Second, we propose a matrix-valued edge weight in DiscoGraph. In such a matrix, each element reflects the inter-agent attention at a specific spatial region, allowing an agent to adaptively highlight the informative regions. During inference, we only need to use the student model named as the distilled collaboration network (DiscoNet). Attributed to the teacher-student framework, multiple agents with the shared DiscoNet could collaboratively approach the performance of a hypothetical teacher model with a holistic view. Our approach is validated on V2X-Sim 1.0, a large-scale multi-agent perception dataset that we synthesized using CARLA and SUMO co-simulation. Our quantitative and qualitative experiments in multi-agent 3D object detection show that DiscoNet could not only achieve a better performance-bandwidth trade-off than the state-of-the-art collaborative perception methods, but also bring more straightforward design rationale. Our code is available on https://github.com/ai4ce/DiscoNet.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 749e48ba-ca6d-4ce6-99e7-fe09a89e229aCited by top-tier papers74
- Where2comm: Communication-Efficient Collaborative Perception via Spatial Confidence MapsYue Hu, Shaoheng Fang, Zixing Lei, Yiqi Zhong et al.NeurIPS 2022 · 537 citations
- DAIR-V2X: A Large-Scale Dataset for Vehicle-Infrastructure Cooperative 3D Object DetectionHaibao Yu, Yizhen Luo, Mao Shu, Yiyi Huo et al.CVPR 2022 · 475 citations
- How2comm: Communication-Efficient and Collaboration-Pragmatic Multi-Agent PerceptionDingkang Yang, Kun Yang, Yuzheng Wang, Jing Liu et al.NeurIPS 2023 · 160 citations
- DeepAccident: A Motion and Accident Prediction Benchmark for V2X Autonomous DrivingTianqi Wang, Sukmin Kim, Wenxuan Ji, Enze Xie et al.AAAI 2024 · 132 citations
- An Extensible Framework for Open Heterogeneous Collaborative PerceptionYifan Lu, Yue Hu, Yiqi Zhong, Dequan Wang et al.ICLR 2024 · 116 citations
Builds on5
- Uncertainty-Aware Multi-Shot Knowledge Distillation for Image-Based Object Re-IdentificationXin Jin, Cuiling Lan, Wenjun Zeng, Zhibo ChenAAAI 2020 · 122 citations
- Fooling LiDAR Perception via Adversarial Trajectory PerturbationYiming Li, Congcong Wen, Felix Juefei-Xu, Chen FengICCV 2021 · 69 citations
- nuScenes: A Multimodal Dataset for Autonomous DrivingHolger Caesar, Varun Bankiti, Alex H. Lang, Sourabh Vora et al.CVPR 2020
- When2com: Multi-Agent Perception via Communication Graph GroupingYen-Cheng Liu, Junjiao Tian, Nathaniel Glaser, Zsolt KiraCVPR 2020
- Point-GNN: Graph Neural Network for 3D Object Detection in a Point CloudWeijing Shi, Raj RajkumarCVPR 2020
Related papers
- UMC: A Unified Bandwidth-efficient and Multi-resolution based Collaborative Perception FrameworkTianhang Wang, Guang Chen, Kai Chen, Zhengfa Liu et al.ICCV 2023 · 46 citations
- Multi-Agent Collaborative Perception via Motion-Aware Robust Communication NetworkShixin Hong, Yu Liu, Zhi Li, Shaohui Li et al.CVPR 2024
- Core: Cooperative Reconstruction for Multi-Agent PerceptionBinglu Wang, Lei Zhang, Zhaozhong Wang, Yongqiang Zhao et al.ICCV 2023 · 73 citations
- DUSA: Decoupled Unsupervised Sim2Real Adaptation for Vehicle-to-Everything Collaborative PerceptionXianghao Kong, Wentao Jiang, Jinrang Jia, Yifeng Shi et al.ACM MM 2023 · 18 citations
- What2comm: Towards Communication-efficient Collaborative Perception via Feature DecouplingKun Yang, Dingkang Yang, Jingyu Zhang, Hanqi Wang et al.ACM MM 2023 · 58 citations
