FocalClick: Towards Practical Interactive Image Segmentation
Xi Chen, Zhiyan Zhao, Yilei Zhang, Manni Duan, Donglian Qi, Hengshuang Zhao
摘要
Interactive segmentation allows users to extract target masks by making positive/negative clicks. Although explored by many previous works, there is still a gap between academic approaches and industrial needs: first, existing models are not efficient enough to work on low-power devices; second, they perform poorly when used to refine preexisting masks as they could not avoid destroying the correct part. FocalClick solves both issues at once by predicting and updating the mask in localized areas. For higher efficiency, we decompose the slow prediction on the entire image into two fast inferences on small crops: a coarse segmentation on the Target Crop, and a local refinement on the Focus Crop. To make the model work with preexisting masks, we formulate a sub-task termed Inter-active Mask Correction, and propose Progressive Merge as the solution. Progressive Merge exploits morphological information to decide where to preserve and where to update, enabling users to refine any preexisting mask effectively. FocalClick achieves competitive results against SOTA methods with significantly smaller FLOPs. It also shows significant superiority when making corrections on preexisting masks. Code and data will be released at github.com/XavierCHEN34/ClickSEG
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper43
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao 等ICCV 2023 · 被引用 13,211 次
- Segment Everything Everywhere All at OnceXueyan Zou, Jianwei Yang, Hao Zhang, Feng Li 等NeurIPS 2023 · 被引用 889 次
- Segment Anything in 3D with NeRFsJiazhong Cen, Zanwei Zhou, Jiemin Fang, Chen Yang 等NeurIPS 2023 · 被引用 255 次
- Segment Any 3D GaussiansJiazhong Cen, Jiemin Fang, Chen Yang, Lingxi Xie 等AAAI 2025 · 被引用 175 次
- SimpleClick: Interactive Image Segmentation with Simple Vision TransformersQin Liu, Zhenlin Xu, Gedas Bertasius, Marc NiethammerICCV 2023 · 被引用 161 次
它引用的顶会 Paper6
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar 等NeurIPS 2021 · 被引用 9,661 次
- Conditional Diffusion for Interactive SegmentationXi Chen, Zhiyan Zhao, Feiwu Yu, Yilei Zhang 等ICCV 2021 · 被引用 100 次
- MultiSeg: Semantically Meaningful, Scale-Diverse Segmentations From Minimal User InputJun Hao Liew, Scott Cohen, Brian L. Price, Long Mai 等ICCV 2019 · 被引用 39 次
- DoveNet: Deep Image Harmonization via Domain VerificationWenyan Cong, Jianfu Zhang, Li Niu, Liu Liu 等CVPR 2020
- F-BRS: Rethinking Backpropagating Refinement for Interactive SegmentationKonstantin Sofiiuk, Ilia A. Petrov, Olga Barinova, Anton KonushinCVPR 2020
相关 Paper
- Efficient Mask Correction for Click-Based Interactive Image SegmentationFei Du, Jianlong Yuan, Zhibin Wang, Fan WangCVPR 2023
- FocusCut: Diving into a Focus View in Interactive SegmentationZheng Lin, Zheng-Peng Duan, Zhao Zhang, Chun-Le Guo 等CVPR 2022 · 被引用 61 次
- Interactive Segmentation with Elaborate Focus PriorKangpeng Hu, Yinghui Sun, Tao Wang, Weihao Zhang 等ICML 2026
- NTClick: Achieving Precise Interactive Segmentation With Noise-tolerant ClicksChenyi Zhang, Ting Liu, Xiaochao Qu, Luoqi Liu 等CVPR 2025
- Focused and Collaborative Feedback Integration for Interactive Image SegmentationQiaoqiao Wei, Hui Zhang, Jun-Hai YongCVPR 2023
