PointConvFormer: Revenge of the Point-based Convolution
Wenxuan Wu, Fuxin Li, Qi Shan
Abstract
We introduce PointConvFormer, a novel building block for point cloud based deep network architectures. Inspired by generalization theory, PointConvFormer combines ideas from point convolution, where filter weights are only based on relative position, and Transformers which utilize featurebased attention. In PointConvFormer, attention computed from feature difference between points in the neighborhood is used to modify the convolutional weights at each point. Hence, we preserved the invariances from point convolution, whereas attention helps to select relevant points in the neighborhood for convolution. PointConvFormer is suitable for multiple tasks that require details at the point level, such as segmentation and scene flow estimation tasks. We experiment on both tasks with multiple datasets including Scan-Net, SemanticKitti, FlyingThings3D and KITTI. Our results show that PointConvFormer offers a better accuracyspeed tradeoff than classic convolutions, regular transformers, and voxelized sparse convolution approaches. Visualizations show that PointConvFormer performs similarly to convolution on flat areas, whereas the neighborhood selection effect is stronger on object boundaries, showing that it has got the best of both worlds. The code will be available.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0da74dff-48dd-4836-891b-0e184919822cCited by top-tier papers11
- Mask-Attention-Free Transformer for 3D Instance SegmentationXin Lai, Yuhui Yuan, Ruihang Chu, Yukang Chen et al.ICCV 2023 · 53 citations
- Towards Large-Scale 3D Representation Learning with Multi-Dataset Point Prompt TrainingXiaoyang Wu, Zhuotao Tian, Xin Wen, Bohao Peng et al.CVPR 2024 · 39 citations
- LitePT: Lighter Yet Stronger Point TransformerYuanwen Yue, Damien Robert, Jianyuan Wang, Sunghwan Hong et al.CVPR 2026 · 25 citations
- KPConvX: Modernizing Kernel Point Convolution with Kernel AttentionHugues Thomas, Yao-Hung Hubert Tsai, Timothy D. Barfoot, Jian ZhangCVPR 2024 · 17 citations
- Spherical Frustum Sparse Convolution Network for LiDAR Point Cloud Semantic SegmentationYu Zheng, Guangming Wang, Jiuming Liu, Marc Pollefeys et al.NeurIPS 2024 · 11 citations
Builds on23
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- KPConv: Flexible and Deformable Convolution for Point CloudsHugues Thomas, Charles R. Qi, Jean-Emmanuel Deschaud, Beatriz Marcotegui et al.ICCV 2019 · 3,193 citations
- SemanticKITTI: A Dataset for Semantic Scene Understanding of LiDAR SequencesJens Behley, Martin Garbade, Andres Milioto, Jan Quenzel et al.ICCV 2019 · 2,345 citations
- DeepGCNs: Can GCNs Go As Deep As CNNs?Guohao Li, Matthias Müller, Ali K. Thabet, Bernard GhanemICCV 2019 · 1,586 citations
- SOLOv2: Dynamic and Fast Instance SegmentationXinlong Wang, Rufeng Zhang, Tao Kong, Lei Li et al.NeurIPS 2020 · 1,193 citations
Related papers
- RPPformer-Flow: Relative Position Guided Point Transformer for Scene Flow EstimationHanlin Li, Guanting Dong, Yueyi Zhang, Xiaoyan Sun et al.ACM MM 2022 · 7 citations
- SCTN: Sparse Convolution-Transformer Network for Scene Flow EstimationBing Li, Cheng Zheng, Silvio Giancola, Bernard GhanemAAAI 2022 · 50 citations
- OctFormer: Octree-based Transformers for 3D Point CloudsPeng-Shuai WangSIGGRAPH 2023 · 123 citations
- GMSF: Global Matching Scene FlowYushan Zhang, Johan Edstedt, Bastian Wandt, Per-Erik Forssén et al.NeurIPS 2023 · 27 citations
- Cloud Transformers: A Universal Approach To Point Cloud Processing TasksKirill Mazur, Victor LempitskyICCV 2021 · 51 citations
