ShapeMask: Learning to Segment Novel Objects by Refining Shape Priors
Weicheng Kuo, Anelia Angelova, Jitendra Malik, Tsung-Yi Lin
摘要
Instance segmentation aims to detect and segment individual objects in a scene. Most existing methods rely on precise mask annotations of every category. However, it is difficult and costly to segment objects in novel categories because a large number of mask annotations is required. We introduce ShapeMask, which learns the intermediate concept of object shape to address the problem of generalization in instance segmentation to novel categories. ShapeMask starts with a bounding box detection and gradually refines it by first estimating the shape of the detected object through a collection of shape priors. Next, ShapeMask refines the coarse shape into an instance level mask by learning instance embeddings. The shape priors provide a strong cue for object-like prediction, and the instance embeddings model the instance specific appearance information. ShapeMask significantly outperforms the state-of-the-art by 6.4 and 3.8 AP when learning across categories, and obtains competitive performance in the fully supervised setting. It is also robust to inaccurate detections, decreased model capacity, and small training data. Moreover, it runs efficiently with 150ms inference time on a GPU and trains within 11 hours on TPUs. With a larger backbone model, ShapeMask increases the gap with state-of-the-art to 9.4 and 6.2 AP across categories. Code will be publicly available at: https://sites.google.com/view/shapemask/home.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- SMIL: Multimodal Learning with Severely Missing ModalityMengmeng Ma, Jian Ren, Long Zhao, Sergey Tulyakov 等AAAI 2021 · 被引用 393 次
- ACFNet: Attentional Class Feature Network for Semantic SegmentationFan Zhang, Yanqin Chen, Zhihang Li, Zhibin Hong 等ICCV 2019 · 被引用 297 次
- Open-Vocabulary Instance Segmentation via Robust Cross-Modal Pseudo-LabelingDat Huynh, Jason Kuen, Zhe Lin, Jiuxiang Gu 等CVPR 2022 · 被引用 78 次
- Weakly-Supervised Text Instance SegmentationXinyan Zu, Haiyang Yu, Bin Li, Xiangyang XueACM MM 2023 · 被引用 8 次
- How to Save your Annotation Cost for Panoptic Segmentation?Xuefeng Du, Chenhan Jiang, Hang Xu, Gengwei Zhang 等AAAI 2021 · 被引用 5 次
相关 Paper
- ContrastMask: Contrastive Learning to Segment Every ThingXuehui Wang, Kai Zhao, Ruixin Zhang, Shouhong Ding 等CVPR 2022 · 被引用 45 次
- Deeply Shape-Guided Cascade for Instance SegmentationHao Ding, Siyuan Qiao, Alan L. Yuille, Wei ShenCVPR 2021
- Learning Saliency Propagation for Semi-Supervised Instance SegmentationYanzhao Zhou, Xin Wang, Jianbin Jiao, Trevor Darrell 等CVPR 2020
- Explicit Shape Encoding for Real-Time Instance SegmentationWenqiang Xu, Haiyang Wang, Fubo Qi, Cewu LuICCV 2019 · 被引用 111 次
- CenterMask: Single Shot Instance Segmentation With Point RepresentationYuqing Wang, Zhaoliang Xu, Hao Shen, Baoshan Cheng 等CVPR 2020
