LCM: Locally Constrained Compact Point Cloud Model for Masked Point Modeling
Yaohua Zha, Naiqi Li, Yanzi Wang, Tao Dai, Hang Guo, Bin Chen, Zhi Wang, Zhihao Ouyang, Shu-Tao Xia
摘要
The pre-trained point cloud model based on Masked Point Modeling (MPM) has exhibited substantial improvements across various tasks. However, these models heavily rely on the Transformer, leading to quadratic complexity and limited decoder, hindering their practice application. To address this limitation, we first conduct a comprehensive analysis of existing Transformer-based MPM, emphasizing the idea that redundancy reduction is crucial for point cloud analysis. To this end, we propose a Locally constrained Compact point cloud Model (LCM) consisting of a locally constrained compact encoder and a locally constrained Mamba-based decoder. Our encoder replaces self-attention with our local aggregation layers to achieve an elegant balance between performance and efficiency. Considering the varying information density between masked and unmasked patches in the decoder inputs of MPM, we introduce a locally constrained Mamba-based decoder. This decoder ensures linear complexity while maximizing the perception of point cloud geometry information from unmasked patches with higher information density. Extensive experimental results show that our compact model significantly surpasses existing Transformer-based models in both performance and efficiency, especially our LCM-based Point-MAE model, compared to the Transformer-based model, achieved an improvement of 1.84%, 0.67%, and 0.60% in average accuracy on the three variants of ScanObjectNN while reducing parameters by 88% and computation by 73%. Code is available at https://github.com/zyh16143998882/LCM.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- PointMamba: A Simple State Space Model for Point Cloud AnalysisDingkang Liang, Xin Zhou, Wei Xu, Xingkui Zhu 等NeurIPS 2024 · 被引用 380 次
- Large Point-to-Gaussian Model for Image-to-3D GenerationLongfei Lu, Huachen Gao, Tao Dai, Yaohua Zha 等ACM MM 2024 · 被引用 9 次
- PointTPA: Dynamic Network Parameter Adaptation for 3D Scene UnderstandingSiyuan Liu, Chaoqun Zheng, Xin Zhou, Tianrui Feng 等CVPR 2026 · 被引用 3 次
- HydraMamba: Multi-Head State Space Model for Global Point Cloud LearningKanglin Qu, Pan Gao, Qun Dai, Yuanhao SunACM MM 2025 · 被引用 2 次
- CloudMamba: Grouped Selective State Spaces for Point Cloud AnalysisKanglin Qu, Pan Gao, Qun Dai, Zhanzhi Ye 等AAAI 2026 · 被引用 2 次
它引用的顶会 Paper42
- BEiT: BERT Pre-Training of Image TransformersHangbo Bao, Li Dong, Songhao Piao, Furu WeiICLR 2022 · 被引用 3,632 次
- Efficiently Modeling Long Sequences with Structured State SpacesAlbert Gu, Karan Goel, Christopher RéICLR 2022 · 被引用 3,482 次
- VMamba: Visual State Space ModelYue Liu, Yunjie Tian, Yuzhong Zhao, Hongtian Yu 等NeurIPS 2024 · 被引用 3,199 次
- Vision Mamba: Efficient Visual Representation Learning with Bidirectional State Space ModelLianghui Zhu, Bencheng Liao, Qian Zhang, Xinlong Wang 等ICML 2024 · 被引用 1,725 次
- Deep Hough Voting for 3D Object Detection in Point CloudsCharles R. Qi, Or Litany, Kaiming He, Leonidas J. GuibasICCV 2019 · 被引用 1,467 次
相关 Paper
- Mamba3D: Enhancing Local Features for 3D Point Cloud Analysis via State Space ModelXu Han, Yuan Tang, Zhaoxuan Wang, Xianzhi LiACM MM 2024 · 被引用 86 次
- 3DET-Mamba: Causal Sequence Modelling for End-to-End 3D Object DetectionMingsheng Li, Jiakang Yuan, Sijin Chen, Lin Zhang 等NeurIPS 2024 · 被引用 5 次
- Pamba: Enhancing Global Interaction in Point Clouds via State Space ModelZhuoyuan Li, Yubo Ai, Jiahao Lu, Chuxin Wang 等AAAI 2025 · 被引用 12 次
- ZigzagPointMamba: Spatial-Semantic Mamba for Point Cloud UnderstandingLinshuang Diao, Sensen Song, Yurong Qian, Dayong RenNeurIPS 2025 · 被引用 9 次
- RWKV3D: An RWKV-Based Model with Multiple Training Strategies for Point Cloud AnalysisChenglong Sun, Shijie Pang, Yuzheng Wang, Lizhe QiACM MM 2025
