LidarMultiNet: Towards a Unified Multi-Task Network for LiDAR Perception
Dongqiangzi Ye, Zixiang Zhou, Weijia Chen, Yufei Xie, Yu Wang, Panqu Wang, Hassan Foroosh
Abstract
LiDAR-based 3D object detection, semantic segmentation, and panoptic segmentation are usually implemented in specialized networks with distinctive architectures that are difficult to adapt to each other. This paper presents LidarMulti-Net, a LiDAR-based multi-task network that unifies these three major LiDAR perception tasks. Among its many benefits, a multi-task network can reduce the overall cost by sharing weights and computation among multiple tasks. However, it typically underperforms compared to independently combined single-task models. The proposed LidarMultiNet aims to bridge the performance gap between the multi-task network and multiple single-task networks. At the core of LidarMultiNet is a strong 3D voxel-based encoder-decoder architecture with a Global Context Pooling (GCP) module extracting global contextual features from a LiDAR frame. Task-specific heads are added on top of the network to perform the three LiDAR perception tasks. More tasks can be implemented simply by adding new task-specific heads while introducing little additional cost. A second stage is also proposed to refine the first-stage segmentation and generate accurate panoptic segmentation results. LidarMultiNet is extensively tested on both Waymo Open Dataset and nuScenes dataset, demonstrating for the first time that major LiDAR perception tasks can be unified in a single strong network that is trained end-to-end and achieves state-of-the-art performance. Notably, LidarMultiNet reaches the official 1 st place in the Waymo Open Dataset 3D semantic segmentation challenge 2022 with the highest mIoU and the best accuracy for most of the 22 classes on the test set, using only LiDAR points as input. It also sets the new state-of-the-art for a single model on the Waymo 3D object detection benchmark and three nuScenes benchmarks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f3791de5-16ba-4d74-9d09-4dce5082524cCited by top-tier papers21
- Segment Anything in 3D with NeRFsJiazhong Cen, Zanwei Zhou, Jiemin Fang, Chen Yang et al.NeurIPS 2023 · 255 citations
- PanoOcc: Unified Occupancy Representation for Camera-based 3D Panoptic SegmentationYuqi Wang, Yuntao Chen, Xingyu Liao, Lue Fan et al.CVPR 2024 · 67 citations
- DetZero: Rethinking Offboard 3D Object Detection with Long-term Sequential Point CloudsTao Ma, Xuemeng Yang, Hongbin Zhou, Xin Li et al.ICCV 2023 · 46 citations
- Query-based Temporal Fusion with Explicit Motion for 3D Object DetectionJinghua Hou, Zhe Liu, Dingkang Liang, Zhikang Zou et al.NeurIPS 2023 · 28 citations
- QuadricFormer: Scene as Superquadrics for 3D Semantic Occupancy PredictionSicheng Zuo, Wenzhao Zheng, Xiaoyong Han, Longchao Yang et al.NeurIPS 2025 · 23 citations
Builds on20
- TransFusion: Robust LiDAR-Camera Fusion for 3D Object Detection with TransformersXuyang Bai, Zeyu Hu, Xinge Zhu, Qingqiu Huang et al.CVPR 2022 · 794 citations
- Sparse Single Sweep LiDAR Point Cloud Segmentation via Learning Contextual Shape Priors from Scene CompletionXu Yan, Jiantao Gao, Jie Li, Ruimao Zhang et al.AAAI 2021 · 365 citations
- Focal Sparse Convolutional Networks for 3D Object DetectionYukang Chen, Yanwei Li, Xiangyu Zhang, Jian Sun et al.CVPR 2022 · 293 citations
- Improving 3D Object Detection with Channel-wise TransformerHualian Sheng, Sijia Cai, Yuan Liu, Bing Deng et al.ICCV 2021 · 293 citations
- RangeDet: In Defense of Range View for LiDAR-based 3D Object DetectionLue Fan, Xuan Xiong, Feng Wang, Naiyan Wang et al.ICCV 2021 · 268 citations
Related papers
- UniSeg: A Unified Multi-Modal LiDAR Segmentation Network and the OpenPCSeg CodebaseYouquan Liu, Runnan Chen, Xin Li, Lingdong Kong et al.ICCV 2023 · 94 citations
- A Versatile Multi-View Framework for LiDAR-based 3D Object Detection with Guidance from Panoptic SegmentationHamidreza Fazlali, Yixuan Xu, Yuan Ren, Bingbing LiuCVPR 2022 · 23 citations
- LiDAR-Based Panoptic Segmentation via Dynamic Shifting NetworkFangzhou Hong, Hui Zhou, Xinge Zhu, Hongsheng Li et al.CVPR 2021
- Multi-Space Alignments Towards Universal LiDAR SegmentationYouquan Liu, Lingdong Kong, Xiaoyang Wu, Runnan Chen et al.CVPR 2024
- Panoptic-PolarNet: Proposal-Free LiDAR Point Cloud Panoptic SegmentationZixiang Zhou, Yang Zhang, Hassan ForooshCVPR 2021
