REFINE: Prediction Fusion Network for Panoptic Segmentation
Jiawei Ren, Cunjun Yu, Zhongang Cai, Mingyuan Zhang, Chongsong Chen, Haiyu Zhao, Shuai Yi, Hongsheng Li
Abstract
Panoptic segmentation aims at generating pixel-wise class and instance predictions for each pixel in the input image, which is a challenging task and far more complicated than naively fusing the semantic and instance segmentation results. Prediction fusion is therefore important to achieve accurate panoptic segmentation. In this paper, we present REFINE, pREdiction FusIon NEtwork for panoptic segmentation, to achieve high-quality panoptic segmentation by improving cross-task prediction fusion, and within-task prediction fusion. Our single-model ResNeXt-101 with DCN achieves PQ=51.5 on the COCO dataset, surpassing state-of-the-art performance by a convincing margin and is comparable with ensembled models. Our smaller model with a ResNet-50 backbone achieves PQ=44.9, which is comparable with state-of-the-art methods with larger backbones.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 77fbae3f-3e89-42fd-8c98-2634be36ffedCited by top-tier papers4
- Panoptic, Instance and Semantic Relations: A Relational Context Encoder to Enhance Panoptic SegmentationShubhankar Borse, Hyojin Park, Hong Cai, Debasmit Das et al.CVPR 2022 · 17 citations
- Multi-view Consistent 3D Panoptic Scene UnderstandingXianzhu Liu, Xin Sun, Haozhe Xie, Zonglin Li et al.AAAI 2025 · 6 citations
- Visual Recognition by RequestChufeng Tang, Lingxi Xie, Xiaopeng Zhang, Xiaolin Hu et al.CVPR 2023
- Open-Vocabulary Panoptic Segmentation with Text-to-Image Diffusion ModelsJiarui Xu, Sifei Liu, Arash Vahdat, Wonmin Byeon et al.CVPR 2023
Builds on4
- SOGNet: Scene Overlap Graph Network for Panoptic SegmentationYibo Yang, Hongyang Li, Xia Li, Qijie Zhao et al.AAAI 2020 · 64 citations
- IMP: Instance Mask Projection for High Accuracy Semantic Segmentation of ThingsCheng-Yang Fu, Tamara L. Berg, Alexander C. BergICCV 2019 · 18 citations
- Learning Instance Occlusion for Panoptic SegmentationJustin Lazarow, Kwonjoon Lee, Kunyu Shi, Zhuowen TuCVPR 2020
- Unifying Training and Inference for Panoptic SegmentationQizhu Li, Xiaojuan Qi, Philip H. S. TorrCVPR 2020
Related papers
- Mask DINO: Towards A Unified Transformer-based Framework for Object Detection and SegmentationFeng Li, Hao Zhang, Huaizhe Xu, Shilong Liu et al.CVPR 2023
- K-Net: Towards Unified Image SegmentationWenwei Zhang, Jiangmiao Pang, Kai Chen, Chen Change LoyNeurIPS 2021 · 500 citations
- Fully Convolutional Networks for Panoptic SegmentationYanwei Li, Hengshuang Zhao, Xiaojuan Qi, Liwei Wang et al.CVPR 2021
- MaX-DeepLab: End-to-End Panoptic Segmentation With Mask TransformersHuiyu Wang, Yukun Zhu, Hartwig Adam, Alan L. Yuille et al.CVPR 2021
- Masked-attention Mask Transformer for Universal Image SegmentationBowen Cheng, Ishan Misra, Alexander G. Schwing, Alexander Kirillov et al.CVPR 2022
