4D Panoptic Segmentation as Invariant and Equivariant Field Prediction
Minghan Zhu, Shizhong Han, Maani Ghaffari, Hong Cai, Fatih Porikli, Shubhankar Borse
Abstract
In this paper, we develop rotation-equivariant neural networks for 4D panoptic segmentation. 4D panoptic segmentation is a benchmark task for autonomous driving that requires recognizing semantic classes and object instances on the road based on LiDAR scans, as well as assigning temporally consistent IDs to instances across time. We observe that the driving scenario is symmetric to rotations on the ground plane. Therefore, rotation-equivariance could provide better generalization and more robust feature learning. Specifically, we review the object instance clustering strategies and restate the centerness-based approach and the offset-based approach as the prediction of invariant scalar fields and equivariant vector fields. Other subtasks are also unified from this perspective, and different invariant and equivariant layers are designed to facilitate their predictions. Through evaluation on the standard 4D panoptic segmentation benchmark of SemanticKITTI, we show that our equivariant models achieve higher accuracy with lower computational costs compared to their non-equivariant counterparts. Moreover, our method sets the new state-of-the-art performance and achieves 1st place on the SemanticKITTI 4D Panoptic Segmentation leaderboard.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 29a07985-7df9-42bb-a5bf-2b544a6cdce4Cited by top-tier papers6
- Equivariant Plug-and-Play Image ReconstructionMatthieu Terris, Thomas Moreau, Nelly Pustelnik, Julián TachellaCVPR 2024 · 25 citations
- TASeg: Temporal Aggregation Network for LiDAR Semantic SegmentationXiaopei Wu, Yuenan Hou, Xiaoshui Huang, Binbin Lin et al.CVPR 2024 · 13 citations
- ZOPP: A Framework of Zero-shot Offboard Panoptic Perception for Autonomous DrivingTao Ma, Hongbin Zhou, Qiusheng Huang, Xuemeng Yang et al.NeurIPS 2024 · 8 citations
- Lie Neurons: Adjoint-Equivariant Neural Networks for Semisimple Lie AlgebrasTzu-Yuan Lin, Minghan Zhu, Maani GhaffariICML 2024 · 6 citations
- 4DSegStreamer: Streaming 4D Panoptic Segmentation via Dual ThreadsLing Liu, Jun Tian, Li YiICCV 2025
Builds on21
- KPConv: Flexible and Deformable Convolution for Point CloudsHugues Thomas, Charles R. Qi, Jean-Emmanuel Deschaud, Beatriz Marcotegui et al.ICCV 2019 · 3,193 citations
- SemanticKITTI: A Dataset for Semantic Scene Understanding of LiDAR SequencesJens Behley, Martin Garbade, Andres Milioto, Jan Quenzel et al.ICCV 2019 · 2,345 citations
- E(n) Equivariant Graph Neural NetworksVictor Garcia Satorras, Emiel Hoogeboom, Max WellingICML 2021 · 1,432 citations
- SE(3)-Transformers: 3D Roto-Translation Equivariant Attention NetworksFabian Fuchs, Daniel E. Worrall, Volker Fischer, Max WellingNeurIPS 2020 · 1,025 citations
- Vector Neurons: A General Framework for SO(3)-Equivariant NetworksCongyue Deng, Or Litany, Yueqi Duan, Adrien Poulenard et al.ICCV 2021 · 411 citations
Related papers
- LiDAR-Based Panoptic Segmentation via Dynamic Shifting NetworkFangzhou Hong, Hui Zhou, Xinge Zhu, Hongsheng Li et al.CVPR 2021
- Panoptic-PolarNet: Proposal-Free LiDAR Point Cloud Panoptic SegmentationZixiang Zhou, Yang Zhang, Hassan ForooshCVPR 2021
- Panoptic-PHNet: Towards Real-Time and High-Precision LiDAR Panoptic Segmentation via Clustering Pseudo HeatmapJinke Li, Xiao He, Yang Wen, Yuan Gao et al.CVPR 2022 · 55 citations
- Center Focusing Network for Real-Time LiDAR Panoptic SegmentationXiaoyan Li, Gang Zhang, Boyue Wang, Yongli Hu et al.CVPR 2023
- SVQNet: Sparse Voxel-Adjacent Query Network for 4D Spatio-Temporal LiDAR Semantic SegmentationXuechao Chen, Shuangjie Xu, Xiaoyi Zou, Tongyi Cao et al.ICCV 2023 · 16 citations
