Instance Segmentation with Mask-supervised Polygonal Boundary Transformers
Justin Lazarow, Weijian Xu, Zhuowen Tu
Abstract
In this paper, we present an end-to-end instance segmentation method that regresses a polygonal boundary for each object instance. This sparse, vectorized boundary representation for objects, while attractive in many downstream computer vision tasks, quickly runs into issues of parity that need to be addressed: parity in supervision and parity in performance when compared to existing pixel-based methods. This is due in part to object instances being annotated with ground-truth in the form of polygonal boundaries or segmentation masks, yet being evaluated in a convenient manner using only segmentation masks. Our method, BoundaryFormer, is a Transformer based architecture that directly predicts polygons yet uses instance mask segmentations as the ground-truth supervision for computing the loss. We achieve this by developing an end-to-end differentiable model that solely relies on supervision within the mask space through differentiable rasterization. Boundary-Former matches or surpasses the Mask R-CNN method in terms of instance segmentation quality on both COCO and Cityscapes while exhibiting significantly better transferability across datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e82cdf0b-da99-4866-b2cc-02dddf79d943Cited by top-tier papers18
- DPText-DETR: Towards Better Scene Text Detection with Dynamic Points in TransformerMaoyuan Ye, Jing Zhang, Shanshan Zhao, Juhua Liu et al.AAAI 2023 · 123 citations
- Online Map Vectorization for Autonomous Driving: A Rasterization PerspectiveGongjie Zhang, Jiahao Lin, Shuang Wu, Yilin Song et al.NeurIPS 2023 · 78 citations
- MapTR: Structured Modeling and Learning for Online Vectorized HD Map ConstructionBencheng Liao, Shaoyu Chen, Xinggang Wang, Tianheng Cheng et al.ICLR 2023 · 69 citations
- BoxSnake: Polygonal Instance Segmentation with Box SupervisionRui Yang, Lin Song, Yixiao Ge, Xiu LiICCV 2023 · 38 citations
- Class-incremental Continual Learning for Instance Segmentation with Image-level Weak SupervisionYu-Hsing Hsieh, Guan-Sheng Chen, Shun-Xian Cai, Ting-Yun Wei et al.ICCV 2023 · 16 citations
Builds on10
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- Deformable DETR: Deformable Transformers for End-to-End Object DetectionXizhou Zhu, Weijie Su, Lewei Lu, Bin Li et al.ICLR 2021 · 7,353 citations
- FCOS: Fully Convolutional One-Stage Object DetectionZhi Tian, Chunhua Shen, Hao Chen, Tong HeICCV 2019 · 6,042 citations
- Soft Rasterizer: A Differentiable Renderer for Image-Based 3D ReasoningShichen Liu, Weikai Chen, Tianye Li, Hao LiICCV 2019 · 789 citations
- End to End Trainable Active Contours via Differentiable RenderingShir Gur, Tal Shaharabany, Lior WolfICLR 2020 · 39 citations
Related papers
- PolyFormer: Referring Image Segmentation as Sequential Polygon GenerationJiang Liu, Hui Ding, Zhaowei Cai, Yuting Zhang et al.CVPR 2023
- Look Closer To Segment Better: Boundary Patch Refinement for Instance SegmentationChufeng Tang, Hang Chen, Xiao Li, Jianmin Li et al.CVPR 2021
- Masked-attention Mask Transformer for Universal Image SegmentationBowen Cheng, Ishan Misra, Alexander G. Schwing, Alexander Kirillov et al.CVPR 2022
- Per-Pixel Classification is Not All You Need for Semantic SegmentationBowen Cheng, Alexander G. Schwing, Alexander KirillovNeurIPS 2021 · 2,196 citations
- Mask Transfiner for High-Quality Instance SegmentationLei Ke, Martin Danelljan, Xia Li, Yu-Wing Tai et al.CVPR 2022
