PatchDCT: Patch Refinement for High Quality Instance Segmentation
Qinrou Wen, Jirui Yang, Xue Yang, Kewei Liang
Abstract
High-quality instance segmentation has shown emerging importance in computer vision. Without any refinement, DCT-Mask directly generates high-resolution masks by compressed vectors. To further refine masks obtained by compressed vectors, we propose for the first time a compressed vector based multi-stage refinement framework. However, the vanilla combination does not bring significant gains, because changes in some elements of the DCT vector will affect the prediction of the entire mask. Thus, we propose a simple and novel method named PatchDCT, which separates the mask decoded from a DCT vector into several patches and refines each patch by the designed classifier and regressor. Specifically, the classifier is used to distinguish mixed patches from all patches, and to correct previously mispredicted foreground and background patches. In contrast, the regressor is used for DCT vector prediction of mixed patches, further refining the segmentation quality at boundary locations. Experiments on COCO show that our method achieves 2.0%, 3.2%, 4.5% AP and 3.4%, 5.3%, 7.0% Boundary AP improvements over Mask-RCNN on COCO, LVIS, and Cityscapes, respectively. It also surpasses DCT-Mask by 0.7%, 1.1%, 1.3% AP and 0.9%, 1.7%, 4.2% Boundary AP on COCO, LVIS and Cityscapes. Besides, the performance of PatchDCT is also competitive with other state-of-the-art methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers5
- Segment Anything in High QualityLei Ke, Mingqiao Ye, Martin Danelljan, Yifan Liu et al.NeurIPS 2023 · 709 citations
- High Quality Entity SegmentationLu Qi, Jason Kuen, Tiancheng Shen, Jiuxiang Gu et al.ICCV 2023 · 91 citations
- Parameter-Inverted Image Pyramid NetworksXizhou Zhu, Xue Yang, Zhaokai Wang, Hao Li et al.NeurIPS 2024 · 11 citations
- E-SAM: Training-Free Segment Every Entity ModelWeiming Zhang, Dingwen Xiao, Lei Chen, Lin WangICCV 2025 · 3 citations
- Robust and Consistent Online Video Instance Segmentation via Instance Mask PropagationMiran Heo, Seoung Wug Oh, Seon Joo Kim, Joon-Young LeeAAAI 2025 · 2 citations
Builds on15
- R3Det: Refined Single-Stage Detector with Feature Refinement for Rotating ObjectXue Yang, Junchi Yan, Ziming Feng, Tao HeAAAI 2021 · 1,109 citations
- SCRDet: Towards More Robust Detection for Small, Cluttered and Rotated ObjectsXue Yang, Jirui Yang, Junchi Yan, Yue Zhang et al.ICCV 2019 · 865 citations
- Learning High-Precision Bounding Box for Rotated Object Detection via Kullback-Leibler DivergenceXue Yang, Xiaojiang Yang, Jirui Yang, Qi Ming et al.NeurIPS 2021 · 603 citations
- Rethinking Rotated Object Detection with Gaussian Wasserstein Distance LossXue Yang, Junchi Yan, Qi Ming, Wentao Wang et al.ICML 2021 · 572 citations
- SOLQ: Segmenting Objects by Learning QueriesBin Dong, Fangao Zeng, Tiancai Wang, Xiangyu Zhang et al.NeurIPS 2021 · 143 citations
Related papers
- RefineMask: Towards High-Quality Instance Segmentation With Fine-Grained FeaturesGang Zhang, Xin Lu, Jingru Tan, Jianmin Li et al.CVPR 2021
- Look Closer To Segment Better: Boundary Patch Refinement for Instance SegmentationChufeng Tang, Hang Chen, Xiao Li, Jianmin Li et al.CVPR 2021
- DCT-Mask: Discrete Cosine Transform Mask Representation for Instance SegmentationXing Shen, Jirui Yang, Chunbo Wei, Bing Deng et al.CVPR 2021
- Mask Transfiner for High-Quality Instance SegmentationLei Ke, Martin Danelljan, Xia Li, Yu-Wing Tai et al.CVPR 2022
- D2Det: Towards High Quality Object Detection and Instance SegmentationJiale Cao, Hisham Cholakkal, Rao Muhammad Anwer, Fahad Shahbaz Khan et al.CVPR 2020
