ShapeMask: Learning to Segment Novel Objects by Refining Shape Priors
Weicheng Kuo, Anelia Angelova, Jitendra Malik, Tsung-Yi Lin
Abstract
Instance segmentation aims to detect and segment individual objects in a scene. Most existing methods rely on precise mask annotations of every category. However, it is difficult and costly to segment objects in novel categories because a large number of mask annotations is required. We introduce ShapeMask, which learns the intermediate concept of object shape to address the problem of generalization in instance segmentation to novel categories. ShapeMask starts with a bounding box detection and gradually refines it by first estimating the shape of the detected object through a collection of shape priors. Next, ShapeMask refines the coarse shape into an instance level mask by learning instance embeddings. The shape priors provide a strong cue for object-like prediction, and the instance embeddings model the instance specific appearance information. ShapeMask significantly outperforms the state-of-the-art by 6.4 and 3.8 AP when learning across categories, and obtains competitive performance in the fully supervised setting. It is also robust to inaccurate detections, decreased model capacity, and small training data. Moreover, it runs efficiently with 150ms inference time on a GPU and trains within 11 hours on TPUs. With a larger backbone model, ShapeMask increases the gap with state-of-the-art to 9.4 and 6.2 AP across categories. Code will be publicly available at: https://sites.google.com/view/shapemask/home.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 046e3fc8-a917-498d-bb78-0ff6235bc88aCited by top-tier papers9
- SMIL: Multimodal Learning with Severely Missing ModalityMengmeng Ma, Jian Ren, Long Zhao, Sergey Tulyakov et al.AAAI 2021 · 393 citations
- ACFNet: Attentional Class Feature Network for Semantic SegmentationFan Zhang, Yanqin Chen, Zhihang Li, Zhibin Hong et al.ICCV 2019 · 297 citations
- Open-Vocabulary Instance Segmentation via Robust Cross-Modal Pseudo-LabelingDat Huynh, Jason Kuen, Zhe Lin, Jiuxiang Gu et al.CVPR 2022 · 78 citations
- Weakly-Supervised Text Instance SegmentationXinyan Zu, Haiyang Yu, Bin Li, Xiangyang XueACM MM 2023 · 8 citations
- How to Save your Annotation Cost for Panoptic Segmentation?Xuefeng Du, Chenhan Jiang, Hang Xu, Gengwei Zhang et al.AAAI 2021 · 5 citations
Related papers
- ContrastMask: Contrastive Learning to Segment Every ThingXuehui Wang, Kai Zhao, Ruixin Zhang, Shouhong Ding et al.CVPR 2022 · 45 citations
- Deeply Shape-Guided Cascade for Instance SegmentationHao Ding, Siyuan Qiao, Alan L. Yuille, Wei ShenCVPR 2021
- Learning Saliency Propagation for Semi-Supervised Instance SegmentationYanzhao Zhou, Xin Wang, Jianbin Jiao, Trevor Darrell et al.CVPR 2020
- Explicit Shape Encoding for Real-Time Instance SegmentationWenqiang Xu, Haiyang Wang, Fubo Qi, Cewu LuICCV 2019 · 111 citations
- CenterMask: Single Shot Instance Segmentation With Point RepresentationYuqing Wang, Zhaoliang Xu, Hao Shen, Baoshan Cheng et al.CVPR 2020
