Retro-FPN: Retrospective Feature Pyramid Network for Point Cloud Semantic Segmentation
Peng Xiang, Xin Wen, Yu-Shen Liu, Hui Zhang, Yi Fang, Zhizhong Han
摘要
Learning per-point semantic features from the hierarchical feature pyramid is essential for point cloud semantic segmentation. However, most previous methods suffered from ambiguous region features or failed to refine per-point features effectively, which leads to information loss and ambiguous semantic identification. To resolve this, we propose Retro-FPN to model the per-point feature prediction as an explicit and retrospective refining process, which goes through all the pyramid layers to extract semantic features explicitly for each point. Its key novelty is a retro-transformer for summarizing semantic contexts from the previous layer and accordingly refining the features in the current stage. In this way, the categorization of each point is conditioned on its local semantic pattern. Specifically, the retro-transformer consists of a local cross-attention block and a semantic gate unit. The cross-attention serves to summarize the semantic pattern retrospectively from the previous layer. And the gate unit carefully incorporates the summarized contexts and refines the current semantic features. Retro-FPN is a pluggable neural network that applies to hierarchical decoders. By integrating Retro-FPN with three representative backbones, including both point-based and voxel-based methods, we show that Retro-FPN can significantly improve performance over state-of-the-art backbones. Comprehensive experiments on widely used benchmarks can justify the effectiveness of our design. The source is available at https://github.com/AllenXiangX/Retro-FPN.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Learning a More Continuous Zero Level Set in Unsigned Distance Fields through Level Set ProjectionJunsheng Zhou, Baorui Ma, Shujuan Li, Yu-Shen Liu 等ICCV 2023 · 被引用 49 次
- Omni-Supervised Motion Editing: Balancing Change and Invariance through Positive-Negative LearningZhenwu Shi, Jingyu Gong, Peiwei Wang, Xingzan Wang 等CVPR 2026 · 被引用 4 次
- HydraMamba: Multi-Head State Space Model for Global Point Cloud LearningKanglin Qu, Pan Gao, Qun Dai, Yuanhao SunACM MM 2025 · 被引用 2 次
- DeepLA-Net: Very Deep Local Aggregation Networks for Point Cloud AnalysisZiyin Zeng, Mingyue Dong, Jian Zhou, Huan Qiu 等CVPR 2025
- Preserving Topological and Geometric Embeddings for Point Cloud RecoveryKaiyue Zhou, Zelong Tan, Hongxiao Wang, Ya-li Li 等AAAI 2026
它引用的顶会 Paper47
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- KPConv: Flexible and Deformable Convolution for Point CloudsHugues Thomas, Charles R. Qi, Jean-Emmanuel Deschaud, Beatriz Marcotegui 等ICCV 2019 · 被引用 3,193 次
- SemanticKITTI: A Dataset for Semantic Scene Understanding of LiDAR SequencesJens Behley, Martin Garbade, Andres Milioto, Jan Quenzel 等ICCV 2019 · 被引用 2,345 次
- PointNeXt: Revisiting PointNet++ with Improved Training and Scaling StrategiesGuocheng Qian, Yuchen Li, Houwen Peng, Jinjie Mai 等NeurIPS 2022 · 被引用 1,270 次
- Point Transformer V2: Grouped Vector Attention and Partition-based PoolingXiaoyang Wu, Yixing Lao, Li Jiang, Xihui Liu 等NeurIPS 2022 · 被引用 924 次
相关 Paper
- CRA-PCN: Point Cloud Completion with Intra- and Inter-level Cross-Resolution TransformersYi Rong, Haoran Zhou, Lixin Yuan, Cheng Mei 等AAAI 2024 · 被引用 37 次
- 3D Object Detection With PointformerXuran Pan, Zhuofan Xia, Shiji Song, Li Erran Li 等CVPR 2021
- Pyramid Architecture for Multi-Scale Processing in Point Cloud SegmentationDong Nie, Rui Lan, Ling Wang, Xiaofeng RenCVPR 2022 · 被引用 38 次
- Why Discard if You can Recycle?: A Recycling Max Pooling Module for 3D Point Cloud AnalysisJiajing Chen, Burak Kakillioglu, Huantao Ren, Senem VelipasalarCVPR 2022 · 被引用 20 次
- SemAffiNet: Semantic-Affine Transformation for Point Cloud SegmentationZiyi Wang, Yongming Rao, Xumin Yu, Jie Zhou 等CVPR 2022 · 被引用 17 次
