Clustering based Point Cloud Representation Learning for 3D Analysis
Tuo Feng, Wenguan Wang, Xiaohan Wang, Yi Yang, Qinghua Zheng
Abstract
Point cloud analysis (such as 3D segmentation and detection) is a challenging task, because of not only the irregular geometries of many millions of unordered points, but also the great variations caused by depth, viewpoint, occlusion, etc. Current studies put much focus on the adaption of neural networks to the complex geometries of point clouds, but are blind to a fundamental question: how to learn an appropriate point embedding space that is aware of both discriminative semantics and challenging variations? As a response, we propose a clustering based supervised learning scheme for point cloud analysis. Unlike current de-facto, scene-wise training paradigm, our algorithm conducts within-class clustering on the point embedding space for automatically discovering subclass patterns which are latent yet representative across scenes. The mined patterns are, in turn, used to repaint the embedding space, so as to respect the underlying distribution of the entire training dataset and improve the robustness to the variations. Our algorithm is principled and readily pluggable to modern point cloud segmentation networks during training, without extra overhead during testing. With various 3D network architectures (i.e., voxel-based, point-based, Transformer-based, automatically searched), our algorithm shows notable improvements on famous point cloud segmentation datasets (i.e., 2.0-2.6% on single-scan and 2.0-2.2% multi-scan of SemanticKITTI, 1.8-1.9% on S3DIS, in terms of mIoU). Our algorithm also demonstrates utility in 3D detection, showing 2.0-3.4% mAP gains on KITTI.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext fd7e3e4a-1244-4af8-a8f8-926b0ca1834eCited by top-tier papers16
- IS-Fusion: Instance-Scene Collaborative Fusion for Multimodal 3D Object DetectionJunbo Yin, Jianbing Shen, Runnan Chen, Wei Li et al.CVPR 2024 · 73 citations
- LogicSeg: Parsing Visual Semantics with Neural Logic Learning and ReasoningLiulei Li, Wenguan Wang, Yang YiICCV 2023 · 52 citations
- Interpretable3D: An Ad-Hoc Interpretable Classifier for 3D Point CloudsTuo Feng, Ruijie Quan, Xiaohan Wang, Wenguan Wang et al.AAAI 2024 · 31 citations
- GroupContrast: Semantic-Aware Self-Supervised Representation Learning for 3D UnderstandingChengyao Wang, Li Jiang, Xiaoyang Wu, Zhuotao Tian et al.CVPR 2024 · 18 citations
- Clustering for Protein Representation LearningRuijie Quan, Wenguan Wang, Fan Ma, Hehe Fan et al.CVPR 2024 · 10 citations
Builds on48
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Unsupervised Learning of Visual Features by Contrasting Cluster AssignmentsMathilde Caron, Ishan Misra, Julien Mairal, Priya Goyal et al.NeurIPS 2020 · 5,249 citations
- KPConv: Flexible and Deformable Convolution for Point CloudsHugues Thomas, Charles R. Qi, Jean-Emmanuel Deschaud, Beatriz Marcotegui et al.ICCV 2019 · 3,193 citations
- SemanticKITTI: A Dataset for Semantic Scene Understanding of LiDAR SequencesJens Behley, Martin Garbade, Andres Milioto, Jan Quenzel et al.ICCV 2019 · 2,345 citations
- DeepGCNs: Can GCNs Go As Deep As CNNs?Guohao Li, Matthias Müller, Ali K. Thabet, Bernard GhanemICCV 2019 · 1,586 citations
Related papers
- PointClustering: Unsupervised Point Cloud Pre-training using Transformation Invariance in ClusteringFuchen Long, Ting Yao, Zhaofan Qiu, Lusong Li et al.CVPR 2023
- Unified 3D Segmenter As Prototypical ClassifiersZheyun Qin, Cheng Han, Qifan Wang, Xiushan Nie et al.NeurIPS 2023 · 27 citations
- Point-GCC: Universal Self-supervised 3D Scene Pre-training via Geometry-Color ContrastGuofan Fan, Zekun Qi, Wenkai Shi, Kaisheng MaACM MM 2024 · 12 citations
- Shape Self-Correction for Unsupervised Point Cloud UnderstandingYe Chen, Jinxian Liu, Bingbing Ni, Hang Wang et al.ICCV 2021 · 58 citations
- Point Cloud Instance Segmentation Using Probabilistic EmbeddingsBiao Zhang, Peter WonkaCVPR 2021
