Interactive Object Segmentation With Inside-Outside Guidance
Shiyin Zhang, Jun Hao Liew, Yunchao Wei, Shikui Wei, Yao Zhao
Abstract
This work explores how to harvest precise object segmentation masks while minimizing the human interaction cost. To achieve this, we propose a simple yet effective interaction scheme, named Inside-Outside Guidance (IOG). Concretely, we leverage an inside point that is clicked near the object center and two outside points at the symmetrical corner locations (top-left and bottom-right or top-right and bottom-left) of an almost-tight bounding box that encloses the target object. The interaction results in a total of one foreground click and four background clicks for segmentation. The advantages of our IOG are four-fold: 1) the two outside points can help remove distractions from other objects or background; 2) the inside point can help eliminate the unrelated regions inside the bounding box; 3) the inside and outside points are easily identified, reducing the confusion raised by the state-of-the-art DEXTR [1] in labeling some extreme samples; 4) it naturally supports additional click annotations for further correction. Despite its simplicity, our IOG not only achieves state-of-the-art performance on several popular benchmarks such as GrabCut [2], PASCAL [3] and MS COCO [4], but also demonstrates strong generalization capability across different domains such as street scenes (Cityscapes [5]), aerial imagery (Rooftop [6] and Agriculture-Vision [7]) and medical images (ssTEM [8]). Code is available at https://github.com/shiyinzhang/Inside-Outside-Guidance .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bf52d0da-db9a-46b9-9145-1d3e31c6aa41Cited by top-tier papers22
- Few-Shot Segmentation via Cycle-Consistent TransformerGengwei Zhang, Guoliang Kang, Yi Yang, Yunchao WeiNeurIPS 2021 · 282 citations
- SimpleClick: Interactive Image Segmentation with Simple Vision TransformersQin Liu, Zhenlin Xu, Gedas Bertasius, Marc NiethammerICCV 2023 · 161 citations
- FocusCut: Diving into a Focus View in Interactive SegmentationZheng Lin, Zheng-Peng Duan, Zhao Zhang, Chun-Le Guo et al.CVPR 2022 · 61 citations
- InterFormer Real-time Interactive Image SegmentationYou Huang, Hao Yang, Ke Sun, Shengchuan Zhang et al.ICCV 2023 · 36 citations
- Interactive Multi-Class Tiny-Object DetectionChunggi Lee, Seonwook Park, Heon Song, Jeongun Ryu et al.CVPR 2022 · 27 citations
Builds on13
- CCNet: Criss-Cross Attention for Semantic SegmentationZilong Huang, Xinggang Wang, Lichao Huang, Chang Huang et al.ICCV 2019 · 2,972 citations
- YOLACT: Real-Time Instance SegmentationDaniel Bolya, Chong Zhou, Fanyi Xiao, Yong Jae LeeICCV 2019 · 2,075 citations
- TensorMask: A Foundation for Dense Object SegmentationXinlei Chen, Ross B. Girshick, Kaiming He, Piotr DollárICCV 2019 · 357 citations
- Integral Object Mining via Online Attention AccumulationPeng-Tao Jiang, Qibin Hou, Yang Cao, Ming-Ming Cheng et al.ICCV 2019 · 246 citations
- SPGNet: Semantic Prediction Guidance for Scene ParsingBowen Cheng, Liang-Chieh Chen, Yunchao Wei, Yukun Zhu et al.ICCV 2019 · 117 citations
Related papers
- Extreme Point Supervised Instance SegmentationHyeonjun Lee, Sehyun Hwang, Suha KwakCVPR 2024
- NTClick: Achieving Precise Interactive Segmentation With Noise-tolerant ClicksChenyi Zhang, Ting Liu, Xiaochao Qu, Luoqi Liu et al.CVPR 2025
- CFR-ICL: Cascade-Forward Refinement with Iterative Click Loss for Interactive Image SegmentationShoukun Sun, Min Xian, Fei Xu, Luca Capriotti et al.AAAI 2024 · 34 citations
- Deeply Shape-Guided Cascade for Instance SegmentationHao Ding, Siyuan Qiao, Alan L. Yuille, Wei ShenCVPR 2021
- Focused and Collaborative Feedback Integration for Interactive Image SegmentationQiaoqiao Wei, Hui Zhang, Jun-Hai YongCVPR 2023
