Point2Mask: Point-supervised Panoptic Segmentation via Optimal Transport
Wentong Li, Yuqian Yuan, Song Wang, Jianke Zhu, Jianshu Li, Jian Liu, Lei Zhang
Abstract
Weakly-supervised image segmentation has recently attracted increasing research attentions, aiming to avoid the expensive pixel-wise labeling. In this paper, we present an effective method, namely Point2Mask, to achieve high-quality panoptic prediction using only a single random point annotation per target for training. Specifically, we formulate the panoptic pseudo-mask generation as an Optimal Transport (OT) problem, where each ground-truth (gt) point label and pixel sample are defined as the label supplier and consumer, respectively. The transportation cost is calculated by the introduced task-oriented maps, which focus on the category-wise and instance-wise differences among the various thing and stuff targets. Furthermore, a centroid-based scheme is proposed to set the accurate unit number for each gt point supplier. Hence, the pseudo-mask generation is converted into finding the optimal transport plan at a globally minimal transportation cost, which can be solved via the Sinkhorn-Knopp Iteration. Experimental results on Pascal VOC and COCO demonstrate the promising performance of our proposed Point2Mask approach to point-supervised panoptic segmentation. Source code is available at: https://github.com/LiWentomng/Point2Mask.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 56fc4809-95b1-49a2-8a77-ed68a6080a78Cited by top-tier papers14
- PointOBB: Learning Oriented Object Detection via Single Point SupervisionJunwei Luo, Xue Yang, Yi Yu, Qingyun Li et al.CVPR 2024 · 41 citations
- Point2RBox: Combine Knowledge from Synthetic Visual Patterns for End-to-End Oriented Object Detection with Single Point SupervisionYi Yu, Xue Yang, Qingyun Li, Feipeng Da et al.CVPR 2024 · 32 citations
- Cam4DOcc: Benchmark for Camera-Only 4D Occupancy Forecasting in Autonomous Driving ApplicationsJunyi Ma, Xieyuanli Chen, Jiawei Huang, Jingyi Xu et al.CVPR 2024 · 28 citations
- Semantic-aware SAM for Point-Prompted Instance SegmentationZhaoyang Wei, Pengfei Chen, Xuehui Yu, Guorong Li et al.CVPR 2024 · 22 citations
- Label-efficient Segmentation via Affinity PropagationWentong Li, Yuqian Yuan, Song Wang, Wenyu Liu et al.NeurIPS 2023 · 11 citations
Builds on17
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- Per-Pixel Classification is Not All You Need for Semantic SegmentationBowen Cheng, Alexander G. Schwing, Alexander KirillovNeurIPS 2021 · 2,196 citations
- K-Net: Towards Unified Image SegmentationWenwei Zhang, Jiangmiao Pang, Kai Chen, Chen Change LoyNeurIPS 2021 · 500 citations
- AdaptIS: Adaptive Instance Selection NetworkKonstantin Sofiiuk, Olga Barinova, Anton KonushinICCV 2019 · 179 citations
- Panoptic SegFormer: Delving Deeper into Panoptic Segmentation with TransformersZhiqi Li, Wenhai Wang, Enze Xie, Zhiding Yu et al.CVPR 2022 · 145 citations
Related papers
- Toward Joint Thing-and-Stuff Mining for Weakly Supervised Panoptic SegmentationYunhang Shen, Liujuan Cao, Zhiwei Chen, Feihong Lian et al.CVPR 2021
- Unsupervised Universal Image SegmentationDantong Niu, Xudong Wang, Xinyang Han, Long Lian et al.CVPR 2024 · 29 citations
- Optimal Transport Minimization: Crowd Localization on Density Maps for Semi-Supervised CountingWei Lin, Antoni B. ChanCVPR 2023
- Fully Data-Driven Pseudo Label Estimation for Pointly-Supervised Panoptic SegmentationJing Li, Junsong Fan, Yuran Yang, Shuqi Mei et al.AAAI 2024 · 2 citations
- OTA: Optimal Transport Assignment for Object DetectionZheng Ge, Songtao Liu, Zeming Li, Osamu Yoshie et al.CVPR 2021
