CFR-ICL: Cascade-Forward Refinement with Iterative Click Loss for Interactive Image Segmentation
Shoukun Sun, Min Xian, Fei Xu, Luca Capriotti, Tiankai Yao
Abstract
The click-based interactive segmentation aims to extract the object of interest from an image with the guidance of user clicks. Recent work has achieved great overall performance by employing feedback from the output. However, in most state-of-the-art approaches, 1) the inference stage involves inflexible heuristic rules and requires a separate refinement model, and 2) the number of user clicks and model performance cannot be balanced. To address the challenges, we propose a click-based and mask-guided interactive image segmentation framework containing three novel components: Cascade-Forward Refinement (CFR), Iterative Click Loss (ICL), and SUEM image augmentation. The CFR offers a unified inference framework to generate segmentation results in a coarse-to-fine manner. The proposed ICL allows model training to improve segmentation and reduce user interactions simultaneously. The proposed SUEM augmentation is a comprehensive way to create large and diverse training sets for interactive image segmentation. Extensive experiments demonstrate the state-of-the-art performance of the proposed approach on five public datasets. Remarkably, our model reduces by 33.2%, and 15.5% the number of clicks required to surpass an IoU of 0.95 in the previous state-of-the-art approach on the Berkeley and DAVIS sets, respectively.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 337fbf44-086b-4200-9064-30d663c8e132Cited by top-tier papers5
- TETRIS: Towards Exploring the Robustness of Interactive SegmentationAndrey Moskalenko, Vlad Shakhuro, Anna Vorontsova, Anton Konushin et al.AAAI 2024 · 3 citations
- DC-TTA: Divide-and-Conquer Framework for Test-Time Adaptation of Interactive SegmentationJihun Kim, Hoyong Kwon, Hyeokjun Kweon, Wooseong Jeong et al.ICCV 2025 · 1 citation
- COCONut: Modernizing COCO SegmentationXueqing Deng, Qihang Yu, Peng Wang, Xiaohui Shen et al.CVPR 2024
- Interactive Segmentation with Elaborate Focus PriorKangpeng Hu, Yinghui Sun, Tao Wang, Weihao Zhang et al.ICML 2026
- Learning and Aligning Click-Aware Shape Prior for Interactive Amodal Instance SegmentationJunjie Chen, Junwei Lin, Ren Hong, Shengjie Liu et al.CVPR 2026
Builds on11
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao et al.ICCV 2023 · 13,211 citations
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar et al.NeurIPS 2021 · 9,661 citations
- AdaptIS: Adaptive Instance Selection NetworkKonstantin Sofiiuk, Olga Barinova, Anton KonushinICCV 2019 · 179 citations
- SimpleClick: Interactive Image Segmentation with Simple Vision TransformersQin Liu, Zhenlin Xu, Gedas Bertasius, Marc NiethammerICCV 2023 · 161 citations
Related papers
- Focused and Collaborative Feedback Integration for Interactive Image SegmentationQiaoqiao Wei, Hui Zhang, Jun-Hai YongCVPR 2023
- Efficient Mask Correction for Click-Based Interactive Image SegmentationFei Du, Jianlong Yuan, Zhibin Wang, Fan WangCVPR 2023
- A Continual Learning Framework for Uncertainty-Aware Interactive Image SegmentationErvine Zheng, Qi Yu, Rui Li, Pengcheng Shi et al.AAAI 2021 · 27 citations
- Interactive Image Segmentation With First Click AttentionZheng Lin, Zhao Zhang, Lin-Zhuo Chen, Ming-Ming Cheng et al.CVPR 2020
- MultiSeg: Semantically Meaningful, Scale-Diverse Segmentations From Minimal User InputJun Hao Liew, Scott Cohen, Brian L. Price, Long Mai et al.ICCV 2019 · 39 citations
