Group Contextual Encoding for 3D Point Clouds
Xu Liu, Chengtao Li, Jian Wang, Jingbo Wang, Boxin Shi, Xiaodong He
摘要
Global context is crucial for 3D point cloud scene understanding tasks. In this work, we extend the contextual encoding layer that was originally designed for 2D tasks to 3D point cloud scenarios. The encoding layer learns a set of code words in feature space of the 3D point cloud to characterize the global semantic context, and then based on these code words, the method learns a global contextual descriptor to reweight the feature maps accordingly. Moreover, compared to 2D scenarios, data sparsity becomes a major issue in 3D point cloud scenarios, and the performance of contextual encoding quickly saturates when the number of code words increases. To mitigate this problem, we further propose a group contextual encoding method, which divides the channel into groups and then performs encoding on group-divided feature vectors. This method facilitates learning of global context in grouped subspace for 3D point clouds. We evaluate the effectiveness and generalizability of our method on three widely-studied 3D point cloud tasks. Experimental results have shown that the proposed method outperformed the VoteNet remarkably with 3 mAP on the benchmark of SUN-RGBD, with the metrics of mAP@0.25, and a much greater margin of 6.57 mAP on ScanNet with the metrics of mAP@0.5. Compared to the baseline of PointNet++, the proposed method leads to an accuracy of 86%, outperforming the baseline by 1.5%.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- HyperDet3D: Learning a Scene-conditioned 3D Object DetectorYu Zheng, Yueqi Duan, Jiwen Lu, Jie Zhou 等CVPR 2022 · 被引用 33 次
- Scene-Aware Generative Network for Human Motion SynthesisJingbo Wang, Sijie Yan, Bo Dai, Dahua LinCVPR 2021
它引用的顶会 Paper2
相关 Paper
- MLCVNet: Multi-Level Context VoteNet for 3D Object DetectionQian Xie, Yu-Kun Lai, Jing Wu, Zhoutao Wang 等CVPR 2020
- PointGroup: Dual-Set Point Grouping for 3D Instance SegmentationLi Jiang, Hengshuang Zhao, Shaoshuai Shi, Shu Liu 等CVPR 2020
- SCF-Net: Learning Spatial Contextual Features for Large-Scale Point Cloud SegmentationSiqi Fan, Qiulei Dong, Fenghua Zhu, Yisheng Lv 等CVPR 2021
- Point Transformer V2: Grouped Vector Attention and Partition-based PoolingXiaoyang Wu, Yixing Lao, Li Jiang, Xihui Liu 等NeurIPS 2022 · 被引用 924 次
- V-DETR: DETR with Vertex Relative Position Encoding for 3D Object DetectionYichao Shen, Zigang Geng, Yuhui Yuan, Yutong Lin 等ICLR 2024 · 被引用 46 次
