Retro-FPN: Retrospective Feature Pyramid Network for Point Cloud Semantic Segmentation
Peng Xiang, Xin Wen, Yu-Shen Liu, Hui Zhang, Yi Fang, Zhizhong Han
Abstract
Learning per-point semantic features from the hierarchical feature pyramid is essential for point cloud semantic segmentation. However, most previous methods suffered from ambiguous region features or failed to refine per-point features effectively, which leads to information loss and ambiguous semantic identification. To resolve this, we propose Retro-FPN to model the per-point feature prediction as an explicit and retrospective refining process, which goes through all the pyramid layers to extract semantic features explicitly for each point. Its key novelty is a retro-transformer for summarizing semantic contexts from the previous layer and accordingly refining the features in the current stage. In this way, the categorization of each point is conditioned on its local semantic pattern. Specifically, the retro-transformer consists of a local cross-attention block and a semantic gate unit. The cross-attention serves to summarize the semantic pattern retrospectively from the previous layer. And the gate unit carefully incorporates the summarized contexts and refines the current semantic features. Retro-FPN is a pluggable neural network that applies to hierarchical decoders. By integrating Retro-FPN with three representative backbones, including both point-based and voxel-based methods, we show that Retro-FPN can significantly improve performance over state-of-the-art backbones. Comprehensive experiments on widely used benchmarks can justify the effectiveness of our design. The source is available at https://github.com/AllenXiangX/Retro-FPN.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9a08f1ca-ecef-4094-87b1-796924ee6b12Cited by top-tier papers8
- Learning a More Continuous Zero Level Set in Unsigned Distance Fields through Level Set ProjectionJunsheng Zhou, Baorui Ma, Shujuan Li, Yu-Shen Liu et al.ICCV 2023 · 49 citations
- Omni-Supervised Motion Editing: Balancing Change and Invariance through Positive-Negative LearningZhenwu Shi, Jingyu Gong, Peiwei Wang, Xingzan Wang et al.CVPR 2026 · 4 citations
- HydraMamba: Multi-Head State Space Model for Global Point Cloud LearningKanglin Qu, Pan Gao, Qun Dai, Yuanhao SunACM MM 2025 · 2 citations
- DeepLA-Net: Very Deep Local Aggregation Networks for Point Cloud AnalysisZiyin Zeng, Mingyue Dong, Jian Zhou, Huan Qiu et al.CVPR 2025
- Preserving Topological and Geometric Embeddings for Point Cloud RecoveryKaiyue Zhou, Zelong Tan, Hongxiao Wang, Ya-li Li et al.AAAI 2026
Builds on47
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- KPConv: Flexible and Deformable Convolution for Point CloudsHugues Thomas, Charles R. Qi, Jean-Emmanuel Deschaud, Beatriz Marcotegui et al.ICCV 2019 · 3,193 citations
- SemanticKITTI: A Dataset for Semantic Scene Understanding of LiDAR SequencesJens Behley, Martin Garbade, Andres Milioto, Jan Quenzel et al.ICCV 2019 · 2,345 citations
- PointNeXt: Revisiting PointNet++ with Improved Training and Scaling StrategiesGuocheng Qian, Yuchen Li, Houwen Peng, Jinjie Mai et al.NeurIPS 2022 · 1,270 citations
- Point Transformer V2: Grouped Vector Attention and Partition-based PoolingXiaoyang Wu, Yixing Lao, Li Jiang, Xihui Liu et al.NeurIPS 2022 · 924 citations
Related papers
- CRA-PCN: Point Cloud Completion with Intra- and Inter-level Cross-Resolution TransformersYi Rong, Haoran Zhou, Lixin Yuan, Cheng Mei et al.AAAI 2024 · 37 citations
- 3D Object Detection With PointformerXuran Pan, Zhuofan Xia, Shiji Song, Li Erran Li et al.CVPR 2021
- Pyramid Architecture for Multi-Scale Processing in Point Cloud SegmentationDong Nie, Rui Lan, Ling Wang, Xiaofeng RenCVPR 2022 · 38 citations
- Why Discard if You can Recycle?: A Recycling Max Pooling Module for 3D Point Cloud AnalysisJiajing Chen, Burak Kakillioglu, Huantao Ren, Senem VelipasalarCVPR 2022 · 20 citations
- SemAffiNet: Semantic-Affine Transformation for Point Cloud SegmentationZiyi Wang, Yongming Rao, Xumin Yu, Jie Zhou et al.CVPR 2022 · 17 citations
