What Makes Good Collaborative Views? Contrastive Mutual Information Maximization for Multi-Agent Perception
Wanfang Su, Lixing Chen, Yang Bai, Xi Lin, Gaolei Li, Zhe Qu, Pan Zhou
Abstract
Multi-agent perception (MAP) allows autonomous systems to understand complex environments by interpreting data from multiple sources. This paper investigates intermediate collaboration for MAP with a specific focus on exploring "good" properties of collaborative view (i.e., post-collaboration feature) and its underlying relationship to individual views (i.e., pre-collaboration features), which were treated as an opaque procedure by most existing works. We propose a novel framework named CMiMC (Contrastive Mutual Information Maximization for Collaborative Perception) for intermediate collaboration. The core philosophy of CMiMC is to preserve discriminative information of individual views in the collaborative view by maximizing mutual information between pre- and post-collaboration features while enhancing the efficacy of collaborative views by minimizing the loss function of downstream tasks. In particular, we define multi-view mutual information (MVMI) for intermediate collaboration that evaluates correlations between collaborative views and individual views on both global and local scales. We establish CMiMNet based on multi-view contrastive learning to realize estimation and maximization of MVMI, which assists the training of a collaborative encoder for voxel-level feature fusion. We evaluate CMiMC on V2X-Sim 1.0, and it improves the SOTA average precision by 3.08% and 4.44% at 0.5 and 0.7 IoU (Intersection-over-Union) thresholds, respectively. In addition, CMiMC can reduce communication volume to 1/32 while achieving performance comparable to SOTA. Code and Appendix are released at https://github.com/77SWF/CMiMC.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext da05fafc-8145-46fa-b0d6-2aa4883bbb50Cited by top-tier papers4
- CoDTS: Enhancing Sparsely Supervised Collaborative Perception with a Dual Teacher-Student FrameworkYushan Han, Hui Zhang, Honglei Zhang, Jing Wang et al.AAAI 2025 · 3 citations
- On the Stability and Generalization of Meta-Learning: the Impact of Inner-LevelsWenjun Ding, Jingling Liu, Lixing Chen, Xiu Su et al.NeurIPS 2025 · 2 citations
- CP-Guard: Malicious Agent Detection and Defense in Collaborative Bird's Eye View PerceptionSenkang Hu, Yihang Tao, Guowen Xu, Yiqin Deng et al.AAAI 2025 · 1 citation
- From Discriminative to Generative: A Diffusion-Based Paradigm for Multi-Agent Collaborative PerceptionKexin Gong, Puyi Yao, Guiyang Luo, Quan Yuan et al.AAAI 2026
Builds on7
- Where2comm: Communication-Efficient Collaborative Perception via Spatial Confidence MapsYue Hu, Shaoheng Fang, Zixing Lei, Yiqi Zhong et al.NeurIPS 2022 · 537 citations
- DAIR-V2X: A Large-Scale Dataset for Vehicle-Infrastructure Cooperative 3D Object DetectionHaibao Yu, Yizhen Luo, Mao Shu, Yiyi Huo et al.CVPR 2022 · 475 citations
- Learning Distilled Collaboration Graph for Multi-Agent PerceptionYiming Li, Shunli Ren, Pengxiang Wu, Siheng Chen et al.NeurIPS 2021 · 464 citations
- Coopernaut: End-to-End Driving with Cooperative Perception for Networked VehiclesJiaxun Cui, Hang Qiu, Dian Chen, Peter Stone et al.CVPR 2022 · 109 citations
- Core: Cooperative Reconstruction for Multi-Agent PerceptionBinglu Wang, Lei Zhang, Zhaozhong Wang, Yongqiang Zhao et al.ICCV 2023 · 73 citations
Related papers
- Complementarity-Enhanced and Redundancy-Minimized Collaboration Network for Multi-agent PerceptionGuiyang Luo, Hui Zhang, Quan Yuan, Jinglin LiACM MM 2022 · 47 citations
- Mutual Contrastive Learning for Visual Representation LearningChuanguang Yang, Zhulin An, Linhang Cai, Yongjun XuAAAI 2022 · 95 citations
- Communication-Efficient Collaborative Perception via Information Filling with CodebookYue Hu, Juntong Peng, Sifei Liu, Junhao Ge et al.CVPR 2024
- Learning Multi-Agent Communication with Contrastive LearningYat Long Lo, Biswa Sengupta, Jakob Nicolaus Foerster, Michael NoukhovitchICLR 2024 · 11 citations
- UMC: A Unified Bandwidth-efficient and Multi-resolution based Collaborative Perception FrameworkTianhang Wang, Guang Chen, Kai Chen, Zhengfa Liu et al.ICCV 2023 · 46 citations
