LiDAR-Camera Panoptic Segmentation via Geometry-Consistent and Semantic-Aware Alignment
Zhiwei Zhang, Zhizhong Zhang, Qian Yu, Ran Yi, Yuan Xie, Lizhuang Ma
Abstract
3D panoptic segmentation is a challenging perception task that requires both semantic segmentation and instance segmentation. In this task, we notice that images could provide rich texture, color, and discriminative information, which can complement LiDAR data for evident performance improvement, but their fusion remains a challenging problem. To this end, we propose LCPS, the first LiDAR-Camera Panoptic Segmentation network. In our approach, we conduct LiDAR-Camera fusion in three stages: 1) an Asynchronous Compensation Pixel Alignment (ACPA) module that calibrates the coordinate misalignment caused by asynchronous problems between sensors; 2) a Semantic-Aware Region Alignment (SARA) module that extends the one-to-one point-pixel mapping to one-to-many semantic relations; 3) a Point-to-Voxel feature Propagation (PVP) module that integrates both geometric and semantic fusion information for the entire point cloud. Our fusion strategy improves about 6.9% PQ performance over the LiDAR-only baseline on NuScenes dataset. Extensive quantitative and qualitative experiments further demonstrate the effectiveness of our novel framework. The code will be released at https://github.com/zhangzw12319/lcps.git.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers10
- TASeg: Temporal Aggregation Network for LiDAR Semantic SegmentationXiaopei Wu, Yuenan Hou, Xiaoshui Huang, Binbin Lin et al.CVPR 2024 · 13 citations
- Beyond the Label Itself: Latent Labels Enhance Semi-supervised Point Cloud Panoptic SegmentationYujun Chen, Xin Tan, Zhizhong Zhang, Yanyun Qu et al.AAAI 2024 · 8 citations
- Unsupervised Modality Adaptation with Text-to-Image Diffusion Models for Semantic SegmentationRuihao Xia, Yu Liang, Peng-Tao Jiang, Hao Zhang et al.NeurIPS 2024 · 7 citations
- Gau-Occ: Geometry-Completed Gaussians for Multi-Modal 3D Occupancy PredictionChengxin Lv, Yihui Li, Hongyu Yang, Yunhong WangCVPR 2026 · 3 citations
- SGFormer: Semantic-Geometry Fusion Transformer for Multi-modal 3D Panoptic SegmentationHongqi Yu, Sixian Chan, Xiaolong Zhou, Xiaoqin ZhangAAAI 2025 · 3 citations
Builds on16
- SemanticKITTI: A Dataset for Semantic Scene Understanding of LiDAR SequencesJens Behley, Martin Garbade, Andres Milioto, Jan Quenzel et al.ICCV 2019 · 2,345 citations
- TransFusion: Robust LiDAR-Camera Fusion for 3D Object Detection with TransformersXuyang Bai, Zeyu Hu, Xinge Zhu, Qingqiu Huang et al.CVPR 2022 · 794 citations
- DeepFusion: Lidar-Camera Deep Fusion for Multi-Modal 3D Object DetectionYingwei Li, Adams Wei Yu, Tianjian Meng, Benjamin Caine et al.CVPR 2022 · 508 citations
- Perception-Aware Multi-Sensor Fusion for 3D LiDAR Semantic SegmentationZhuangwei Zhuang, Rong Li, Kui Jia, Qicheng Wang et al.ICCV 2021 · 129 citations
- GP-S3Net: Graph-based Panoptic Sparse Semantic Segmentation NetworkRyan Razani, Ran Cheng, Enxu Li, Ehsan Taghavi et al.ICCV 2021 · 60 citations
Related papers
- UniSeg: A Unified Multi-Modal LiDAR Segmentation Network and the OpenPCSeg CodebaseYouquan Liu, Runnan Chen, Xin Li, Lingdong Kong et al.ICCV 2023 · 94 citations
- LiDAR-Based Panoptic Segmentation via Dynamic Shifting NetworkFangzhou Hong, Hui Zhou, Xinge Zhu, Hongsheng Li et al.CVPR 2021
- How Do Images Align and Complement LiDAR? Towards a Harmonized Multi-modal 3D Panoptic SegmentationYining Pan, Qiongjie Cui, Xulei Yang, Na ZhaoICML 2025
- LidarMultiNet: Towards a Unified Multi-Task Network for LiDAR PerceptionDongqiangzi Ye, Zixiang Zhou, Weijia Chen, Yufei Xie et al.AAAI 2023 · 108 citations
- Panoptic-PolarNet: Proposal-Free LiDAR Point Cloud Panoptic SegmentationZixiang Zhou, Yang Zhang, Hassan ForooshCVPR 2021
