GraCo: Granularity-Controllable Interactive Segmentation
Yian Zhao, Kehan Li, Zesen Cheng, Pengchong Qiao, Xiawu Zheng, Rongrong Ji, Chang Liu, Li Yuan, Jie Chen
Abstract
Interactive Segmentation (IS) segments specific objects or parts in the image according to user input. Current IS pipelines fall into two categories: single-granularity out-put and multi-granularity output. The latter aims to allevi-ate the spatial ambiguity present in the former. However, the multi-granularity output pipeline suffers from limited interaction flexibility and produces redundant results. In this work, we introduce Granularity-Controllable Interactive Segmentation (GraCo), a novel approach that allows precise control of prediction granularity by introducing ad-ditional parameters to input. This enhances the customization of the interactive system and eliminates redundancy while resolving ambiguity. Nevertheless, the exorbitant cost of annotating multi-granularity masks and the lack of avail-able datasets with granularity annotations make it difficult for models to acquire the necessary guidance to control out-put granularity. To address this problem, we design an any-granularity mask generator that exploits the semantic property of the pre-trained IS model to automatically gen-erate abundant mask-granularity pairs without requiring additional manual annotation. Based on these pairs, we propose a granularity-controllable learning strategy that efficiently imparts the granularity controllability to the IS model. Extensive experiments on intricate scenarios at ob-ject and part levels demonstrate that our GraCo has signifi-cant advantages over previous methods. This highlights the potential of GraCo to be a flexible annotation tool, capable of adapting to diverse segmentation scenarios. The project page: https://zhao-yian.github.io/GraCo.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ff9b655b-16ee-411e-bae4-962d62f44274Cited by top-tier papers5
- Tune-Your-Style: Intensity-Tunable 3D Style Transfer with Gaussian SplattingYian Zhao, Rushi Ye, Ruochong Zheng, Zesen Cheng et al.ICCV 2025 · 3 citations
- DC-TTA: Divide-and-Conquer Framework for Test-Time Adaptation of Interactive SegmentationJihun Kim, Hoyong Kwon, Hyeokjun Kweon, Wooseong Jeong et al.ICCV 2025 · 1 citation
- VASparse: Towards Efficient Visual Hallucination Mitigation via Visual-Aware Token SparsificationXianwei Zhuang, Zhihong Zhu, Yuxin Xie, Liming Liang et al.CVPR 2025
- Order-aware Interactive SegmentationBin Wang, Anwesa Choudhuri, Meng Zheng, Zhongpai Gao et al.ICLR 2025
- Learning and Aligning Click-Aware Shape Prior for Interactive Amodal Instance SegmentationJunjie Chen, Junwei Lin, Ren Hong, Shengjie Liu et al.CVPR 2026
Builds on19
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao et al.ICCV 2023 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 6,759 citations
- DETRs Beat YOLOs on Real-time Object DetectionYian Zhao, Wenyu Lv, Shangliang Xu, Jinman Wei et al.CVPR 2024 · 3,046 citations
Related papers
- MultiSeg: Semantically Meaningful, Scale-Diverse Segmentations From Minimal User InputJun Hao Liew, Scott Cohen, Brian L. Price, Long Mai et al.ICCV 2019 · 39 citations
- Multi-granularity Interaction Simulation for Unsupervised Interactive SegmentationKehan Li, Yian Zhao, Zhennan Wang, Zesen Cheng et al.ICCV 2023 · 11 citations
- MuGE: Multiple Granularity Edge DetectionCaixia Zhou, Yaping Huang, Mengyang Pu, Qingji Guan et al.CVPR 2024 · 25 citations
- Variance-Insensitive and Target-Preserving Mask Refinement for Interactive Image SegmentationChaowei Fang, Ziyin Zhou, Junye Chen, Hanjing Su et al.AAAI 2024 · 7 citations
- Universal Segmentation at Arbitrary Granularity with Language InstructionYong Liu, Cairong Zhang, Yitong Wang, Jiahao Wang et al.CVPR 2024 · 15 citations
