Generative Map Priors for Collaborative BEV Semantic Segmentation
Jiahui Fu, Yue Gong, Luting Wang, Shifeng Zhang, Xu Zhou, Si Liu
Abstract
Collaborative perception aims to address the constraint of single-agent perception by exchanging information among multiple agents. Previous works primarily focus on collaborative object detection, exploring compressed transmission and fusion prediction tailored to sparse object features. However, these strategies are not well-suited for dense features in collaborative BEV semantic segmentation. Therefore, we propose CoGMP, a novel Collaborative framework that leverages Generative Map Priors to achieve efficient compression and robust fusion. CoGMP introduces two key innovations: Element Format Feature Compression (EFFC) and Structure Guided Feature Fusion (SGFF). Specifically, EFFC leverages map element priors from codebook to encode BEV features as discrete element indices for transmitted information compression. Meanwhile, SGFF utilizes a diffusion model with structural priors to coherently integrate multi-agent features, thereby achieving consistent fusion predictions. Evaluations on the OPV2V dataset show that CoGMP achieves a 6.89/7.64 Road/Lane IoU improvement and a 32-fold reduction in communication volume.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e7e292d4-a5d6-4c91-9dd7-1790b26f004cCited by top-tier papers5
- Linking Modality Isolation in Heterogeneous Collaborative PerceptionChangxing Liu, Zichen Chao, Siheng ChenCVPR 2026 · 3 citations
- WhisperNet: A Scalable Solution for Bandwidth-Efficient CollaborationGong Chen, Chaokun Zhang, Xinyan ZhaoCVPR 2026 · 3 citations
- CycleBEV: Regularizing View Transformation Networks via View Cycle Consistency for Bird’s-Eye-View Semantic SegmentationJeongbin Hong, Dooseop Choi, Taeg-Hyun An, KYOUNG AN AN et al.CVPR 2026 · 1 citation
- CoopDiff: A Diffusion-Guided Approach for Cooperation under CorruptionsGong Chen, Chaokun Zhang, Pengcheng LvCVPR 2026 · 1 citation
- CoLC: Communication-Efficient Collaborative Perception with LiDAR CompletionYushan Han, Hui Zhang, Qiming Xia, Yi Jin et al.CVPR 2026
Builds on39
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 6,759 citations
- DiffusionDet: Diffusion Model for Object DetectionShoufa Chen, Peize Sun, Yibing Song, Ping LuoICCV 2023 · 715 citations
Related papers
- Communication-Efficient Multi-Vehicle Collaborative Semantic Segmentation via Sparse 3D Gaussian SharingTianyu Hong, Xiaobo Zhou, Wenkai Hu, Qi Xie et al.ICCV 2025 · 2 citations
- Point Cluster: A Compact Message Unit for Communication-Efficient Collaborative PerceptionZihan Ding, Jiahui Fu, Si Liu, Hongyu Li et al.ICLR 2025
- From Discriminative to Generative: A Diffusion-Based Paradigm for Multi-Agent Collaborative PerceptionKexin Gong, Puyi Yao, Guiyang Luo, Quan Yuan et al.AAAI 2026
- Core: Cooperative Reconstruction for Multi-Agent PerceptionBinglu Wang, Lei Zhang, Zhaozhong Wang, Yongqiang Zhao et al.ICCV 2023 · 73 citations
- Collaborative Semantic Occupancy Prediction with Hybrid Feature Fusion in Connected Automated VehiclesRui Song, Chenwei Liang, Hu Cao, Zhiran Yan et al.CVPR 2024
