What Makes Good Collaborative Views? Contrastive Mutual Information Maximization for Multi-Agent Perception
Wanfang Su, Lixing Chen, Yang Bai, Xi Lin, Gaolei Li, Zhe Qu, Pan Zhou
摘要
Multi-agent perception (MAP) allows autonomous systems to understand complex environments by interpreting data from multiple sources. This paper investigates intermediate collaboration for MAP with a specific focus on exploring "good" properties of collaborative view (i.e., post-collaboration feature) and its underlying relationship to individual views (i.e., pre-collaboration features), which were treated as an opaque procedure by most existing works. We propose a novel framework named CMiMC (Contrastive Mutual Information Maximization for Collaborative Perception) for intermediate collaboration. The core philosophy of CMiMC is to preserve discriminative information of individual views in the collaborative view by maximizing mutual information between pre- and post-collaboration features while enhancing the efficacy of collaborative views by minimizing the loss function of downstream tasks. In particular, we define multi-view mutual information (MVMI) for intermediate collaboration that evaluates correlations between collaborative views and individual views on both global and local scales. We establish CMiMNet based on multi-view contrastive learning to realize estimation and maximization of MVMI, which assists the training of a collaborative encoder for voxel-level feature fusion. We evaluate CMiMC on V2X-Sim 1.0, and it improves the SOTA average precision by 3.08% and 4.44% at 0.5 and 0.7 IoU (Intersection-over-Union) thresholds, respectively. In addition, CMiMC can reduce communication volume to 1/32 while achieving performance comparable to SOTA. Code and Appendix are released at https://github.com/77SWF/CMiMC.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- CoDTS: Enhancing Sparsely Supervised Collaborative Perception with a Dual Teacher-Student FrameworkYushan Han, Hui Zhang, Honglei Zhang, Jing Wang 等AAAI 2025 · 被引用 3 次
- On the Stability and Generalization of Meta-Learning: the Impact of Inner-LevelsWenjun Ding, Jingling Liu, Lixing Chen, Xiu Su 等NeurIPS 2025 · 被引用 2 次
- CP-Guard: Malicious Agent Detection and Defense in Collaborative Bird's Eye View PerceptionSenkang Hu, Yihang Tao, Guowen Xu, Yiqin Deng 等AAAI 2025 · 被引用 1 次
- From Discriminative to Generative: A Diffusion-Based Paradigm for Multi-Agent Collaborative PerceptionKexin Gong, Puyi Yao, Guiyang Luo, Quan Yuan 等AAAI 2026
它引用的顶会 Paper7
- Where2comm: Communication-Efficient Collaborative Perception via Spatial Confidence MapsYue Hu, Shaoheng Fang, Zixing Lei, Yiqi Zhong 等NeurIPS 2022 · 被引用 537 次
- DAIR-V2X: A Large-Scale Dataset for Vehicle-Infrastructure Cooperative 3D Object DetectionHaibao Yu, Yizhen Luo, Mao Shu, Yiyi Huo 等CVPR 2022 · 被引用 475 次
- Learning Distilled Collaboration Graph for Multi-Agent PerceptionYiming Li, Shunli Ren, Pengxiang Wu, Siheng Chen 等NeurIPS 2021 · 被引用 464 次
- Coopernaut: End-to-End Driving with Cooperative Perception for Networked VehiclesJiaxun Cui, Hang Qiu, Dian Chen, Peter Stone 等CVPR 2022 · 被引用 109 次
- Core: Cooperative Reconstruction for Multi-Agent PerceptionBinglu Wang, Lei Zhang, Zhaozhong Wang, Yongqiang Zhao 等ICCV 2023 · 被引用 73 次
相关 Paper
- Complementarity-Enhanced and Redundancy-Minimized Collaboration Network for Multi-agent PerceptionGuiyang Luo, Hui Zhang, Quan Yuan, Jinglin LiACM MM 2022 · 被引用 47 次
- Mutual Contrastive Learning for Visual Representation LearningChuanguang Yang, Zhulin An, Linhang Cai, Yongjun XuAAAI 2022 · 被引用 95 次
- Communication-Efficient Collaborative Perception via Information Filling with CodebookYue Hu, Juntong Peng, Sifei Liu, Junhao Ge 等CVPR 2024
- Learning Multi-Agent Communication with Contrastive LearningYat Long Lo, Biswa Sengupta, Jakob Nicolaus Foerster, Michael NoukhovitchICLR 2024 · 被引用 11 次
- UMC: A Unified Bandwidth-efficient and Multi-resolution based Collaborative Perception FrameworkTianhang Wang, Guang Chen, Kai Chen, Zhengfa Liu 等ICCV 2023 · 被引用 46 次
