SemAffiNet: Semantic-Affine Transformation for Point Cloud Segmentation
Ziyi Wang, Yongming Rao, Xumin Yu, Jie Zhou, Jiwen Lu
摘要
Conventional point cloud semantic segmentation methods usually employ an encoder-decoder architecture, where mid-level features are locally aggregated to extract geometric information. However, the over-reliance on these classagnostic local geometric representations may raise confusion between local parts from different categories that are similar in appearance or spatially adjacent. To address this issue, we argue that mid-level features can be further enhanced with semantic information, and propose semanticaffine transformation that transforms features of mid-level points belonging to different categories with class-specific affine parameters. Based on this technique, we propose Se-mAffiNet for point cloud semantic segmentation, which utilizes the attention mechanism in the Transformer module to implicitly and explicitly capture global structural knowledge within local parts for overall comprehension of each category. We conduct extensive experiments on the Scan-NetV2 and NYUv2 datasets, and evaluate semantic-affine transformation on various 3D point cloud and 2D image segmentation baselines, where both qualitative and quantitative results demonstrate the superiority and generalization ability of our proposed approach. Code is available at https://github.com/wangzy22/SemAffiNet .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- 2D-3D Interlaced Transformer for Point Cloud Segmentation with Scene-Level SupervisionCheng-Kun Yang, Min-Hung Chen, Yung-Yu Chuang, Yen-Yu LinICCV 2023 · 被引用 30 次
- MCUFormer: Deploying Vision Tranformers on Microcontrollers with Limited MemoryYinan Liang, Ziwei Wang, Xiuwei Xu, Yansong Tang 等NeurIPS 2023 · 被引用 26 次
- Attention Discriminant Sampling for Point CloudsCheng-Yao Hong, Yu-Ying Chou, Tyng-Luh LiuICCV 2023 · 被引用 21 次
- Generalized Few-Shot Point Cloud Segmentation Via Geometric WordsYating Xu, Conghui Hu, Na Zhao, Gim Hee LeeICCV 2023 · 被引用 18 次
- MirageRoom: 3D Scene Segmentation with 2D Pre-Trained Models by Mirage ProjectionHaowen Sun, Yueqi Duan, Juncheng Yan, Yifan Liu 等CVPR 2024 · 被引用 5 次
它引用的顶会 Paper20
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- KPConv: Flexible and Deformable Convolution for Point CloudsHugues Thomas, Charles R. Qi, Jean-Emmanuel Deschaud, Beatriz Marcotegui 等ICCV 2019 · 被引用 3,193 次
- Per-Pixel Classification is Not All You Need for Semantic SegmentationBowen Cheng, Alexander G. Schwing, Alexander KirillovNeurIPS 2021 · 被引用 2,196 次
- PoinTr: Diverse Point Cloud Completion with Geometry-Aware TransformersXumin Yu, Yongming Rao, Ziyi Wang, Zuyan Liu 等ICCV 2021 · 被引用 592 次
相关 Paper
- Point Transformer V2: Grouped Vector Attention and Partition-based PoolingXiaoyang Wu, Yixing Lao, Li Jiang, Xihui Liu 等NeurIPS 2022 · 被引用 924 次
- RGGT: A Generative-Prior-Guided Transformer for Unified Rigid and Non-Rigid Point Cloud RegistrationChengyu Zheng, Songlin Yang, Jin Huang, Honghua Chen 等ICML 2026
- Unified 3D Segmenter As Prototypical ClassifiersZheyun Qin, Cheng Han, Qifan Wang, Xiushan Nie 等NeurIPS 2023 · 被引用 27 次
- Point TransformerHengshuang Zhao, Li Jiang, Jiaya Jia, Philip H. S. Torr 等ICCV 2021 · 被引用 23 次
- PointConvFormer: Revenge of the Point-based ConvolutionWenxuan Wu, Fuxin Li, Qi ShanCVPR 2023
