Unified 3D Segmenter As Prototypical Classifiers
Zheyun Qin, Cheng Han, Qifan Wang, Xiushan Nie, Yilong Yin, Xiankai Lu
摘要
The task of point cloud segmentation, comprising semantic, instance, and panoptic segmentation, has been mainly tackled by designing task-specific network architectures, which often lack the flexibility to generalize across tasks, thus resulting in a fragmented research landscape. In this paper, we introduce P ROTO SEG, a prototype-based model that unifies semantic, instance, and panoptic segmentation tasks. Our approach treats these three homogeneous tasks as a classification problem with different levels of granularity. By leveraging a Transformer architecture, we extract point embeddings to optimize prototype-class distances and dynamically learn class prototypes to accommodate the end tasks. Our prototypical design enjoys simplicity and transparency, powerful representational learning, and ad-hoc explainability. Empirical results demonstrate that P ROTO SEG outperforms concurrent well-known specialized architectures on 3D point cloud benchmarks, achieving 72.3%, 76.4% and 74.2% mIoU for semantic segmentation on S3DIS, ScanNet V2 and SemanticKITTI, 66.8% mCov and 51.2% mAP for instance segmentation on S3DIS and ScanNet V2, 62.4% PQ for panoptic segmentation on SemanticKITTI, validating the strength of our concept and the effectiveness of our algorithm. The code and models are available at https://github.com/zyqin19/PROTOSEG .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Facing the Elephant in the Room: Visual Prompt Tuning or Full finetuning?Cheng Han, Qifan Wang, Yiming Cui, Wenguan Wang 等ICLR 2024 · 被引用 43 次
- Prototypical Transformer As Unified Motion LearnersCheng Han, Yawen Lu, Guohao Sun, James Chenhao Liang 等ICML 2024 · 被引用 9 次
- Learning Unknowns from Unknowns: Diversified Negative Prototypes Generator for Few-shot Open-Set RecognitionZhenyu Zhang, Guangyao Chen, Yixiong Zou, Yuhua Li 等ACM MM 2024 · 被引用 7 次
- Learning Clustering-based Prototypes for Compositional Zero-Shot LearningHongyu Qu, Jianan Wei, Xiangbo Shu, Wenguan WangICLR 2025
- Semantic and Sequential Alignment for Referring Video Object SegmentationFeiyu Pan, Hao Fang, Fangkai Li, Yanyu Xu 等CVPR 2025
它引用的顶会 Paper46
- Unsupervised Learning of Visual Features by Contrasting Cluster AssignmentsMathilde Caron, Ishan Misra, Julien Mairal, Priya Goyal 等NeurIPS 2020 · 被引用 5,249 次
- SemanticKITTI: A Dataset for Semantic Scene Understanding of LiDAR SequencesJens Behley, Martin Garbade, Andres Milioto, Jan Quenzel 等ICCV 2019 · 被引用 2,345 次
- Per-Pixel Classification is Not All You Need for Semantic SegmentationBowen Cheng, Alexander G. Schwing, Alexander KirillovNeurIPS 2021 · 被引用 2,196 次
- PointNeXt: Revisiting PointNet++ with Improved Training and Scaling StrategiesGuocheng Qian, Yuchen Li, Houwen Peng, Jinjie Mai 等NeurIPS 2022 · 被引用 1,270 次
- Point Transformer V2: Grouped Vector Attention and Partition-based PoolingXiaoyang Wu, Yixing Lao, Li Jiang, Xihui Liu 等NeurIPS 2022 · 被引用 924 次
相关 Paper
- OneFormer3D: One Transformer for Unified Point Cloud SegmentationMaxim Kolodiazhnyi, Anna Vorontsova, Anton Konushin, Danila RukhovichCVPR 2024
- PUPS: Point Cloud Unified Panoptic SegmentationShihao Su, Jianyun Xu, Huanyu Wang, Zhenwei Miao 等AAAI 2023 · 被引用 30 次
- A Unified Framework for 3D Scene UnderstandingWei Xu, Chunsheng Shi, Sifan Tu, Xin Zhou 等NeurIPS 2024 · 被引用 25 次
- PointGroup: Dual-Set Point Grouping for 3D Instance SegmentationLi Jiang, Hengshuang Zhao, Shaoshuai Shi, Shu Liu 等CVPR 2020
- Clustering based Point Cloud Representation Learning for 3D AnalysisTuo Feng, Wenguan Wang, Xiaohan Wang, Yi Yang 等ICCV 2023 · 被引用 53 次
