PointPatchMix: Point Cloud Mixing with Patch Scoring
Yi Wang, Jiaze Wang, Jinpeng Li, Zixu Zhao, Guangyong Chen, Anfeng Liu, Pheng-Ann Heng
Abstract
Data augmentation is an effective regularization strategy for mitigating overfitting in deep neural networks, and it plays a crucial role in 3D vision tasks, where the point cloud data is relatively limited. While mixing-based augmentation has shown promise for point clouds, previous methods mix point clouds either on block level or point level, which has constrained their ability to strike a balance between generating diverse training samples and preserving the local characteristics of point clouds. Additionally, the varying importance of each part of the point clouds has not been fully considered, cause not all parts contribute equally to the classification task, and some parts may contain unimportant or redundant information. To overcome these challenges, we propose PointPatchMix, a novel approach that mixes point clouds at the patch level and integrates a patch scoring module to generate content-based targets for mixed point clouds. Our approach preserves local features at the patch level, while the patch scoring module assigns targets based on the content-based significance score from a pre-trained teacher model. We evaluate PointPatchMix on two benchmark datasets, ModelNet40 and ScanObjectNN, and demonstrate significant improvements over various baselines in both synthetic and realworld datasets, as well as few-shot settings. With Point-MAE as our baseline, our model surpasses previous methods by a significant margin, achieving 86.3% accuracy on ScanObjectNN and 94.1% accuracy on ModelNet40. Furthermore, our approach shows strong generalization across multiple architectures and enhances the robustness of the baseline model.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers5
- MM-Mixing: Multi-Modal Mixing Alignment for 3D UnderstandingJiaze Wang, Yi Wang, Ziyu Guo, Renrui Zhang et al.AAAI 2025 · 1 citation
- What We Miss Matters: Learning from the Overlooked in Point Cloud TransformersYi Wang, Jiaze Wang, Ziyu Guo, Renrui Zhang et al.NeurIPS 2025 · 1 citation
- Harnessing Text-to-Image Diffusion Models for Point Cloud Self-Supervised LearningYiyang Chen, Shanshan Zhao, Lunhao Duan, Changxing Ding et al.ICCV 2025
- Shaping Without Tearing: Controllable Diffeomorphic Deformations for Topology-Preserving 3D Point Cloud AugmentationJian Bi, Qianliang Wu, Jianjun Qian, Lei Luo et al.AAAI 2026
- Exploring Scene Affinity for Semi-Supervised LiDAR Semantic SegmentationChuandong Liu, Xingxing Weng, Shuguo Jiang, Pengcheng Li et al.CVPR 2025
Builds on23
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- CutMix: Regularization Strategy to Train Strong Classifiers With Localizable FeaturesSangdoo Yun, Dongyoon Han, Sanghyuk Chun, Seong Joon Oh et al.ICCV 2019 · 5,843 citations
- Pyramid Vision Transformer: A Versatile Backbone for Dense Prediction without ConvolutionsWenhai Wang, Enze Xie, Xiang Li, Deng-Ping Fan et al.ICCV 2021 · 4,909 citations
- KPConv: Flexible and Deformable Convolution for Point CloudsHugues Thomas, Charles R. Qi, Jean-Emmanuel Deschaud, Beatriz Marcotegui et al.ICCV 2019 · 3,193 citations
Related papers
- SageMix: Saliency-Guided Mixup for Point CloudsSanghyeok Lee, Minkyu Jeon, Injae Kim, Yunyang Xiong et al.NeurIPS 2022 · 37 citations
- Regularization Strategy for Point Cloud via Rigidly Mixed SampleDogyoon Lee, Jaeha Lee, Junhyeop Lee, Hyeongmin Lee et al.CVPR 2021
- Semi-supervised 3D Object Detection with PatchTeacher and PillarMixXiaopei Wu, Liang Peng, Liang Xie, Yuenan Hou et al.AAAI 2024 · 10 citations
- PSMix: Robust Point Cloud Recognition through Spectral Domain MixingXin Wei, Qin Yang, Hongji Zhao, Fei Gao et al.ICML 2026
- Point-BERT: Pre-training 3D Point Cloud Transformers with Masked Point ModelingXumin Yu, Lulu Tang, Yongming Rao, Tiejun Huang et al.CVPR 2022 · 684 citations
