Learning and Aligning Click-Aware Shape Prior for Interactive Amodal Instance Segmentation
Junjie Chen, Junwei Lin, Ren Hong, Shengjie Liu, Yuming Fang, Feng Qian, Yifan Zuo
摘要
Amodal instance segmentation aims to segment both visible and occluded regions of object instance, which are challenging due to lacking inference support under occlusion. Most existing methods employ the prior knowledge about object mask (shape prior) to support the amodal estimation, but the shape prior is not always compatible for object instances in the test stage. In this paper, we explore the task of interactive amodal segmentation, where a few user clicks are available for better segmenting the complete masks of object instances. For this task, we propose a novel framework based on learning and aligning clickaware shape prior (termed ClickPriorNet). Specifically, we propose to learn click-aware shape prior with triplet loss, which forces the retrieved shape priors to have higher IoU with the ground-truth of target instance and thus could exactly facilitate the prediction. Besides, considering the inevitable mismatch between shape prior and target instance, we propose to adaptively align the shape prior with deformable attention. Overall, our model could make full use of the interactive clicks to retrieve and align shape priors, and thus could estimate more complete masks. Extensive experiments on three benchmark datasets demonstrate the effectiveness of our method. Code is available at https: //github.com/chenbys/ClickPriorNet.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper31
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Deformable DETR: Deformable Transformers for End-to-End Object DetectionXizhou Zhu, Weijie Su, Lewei Lu, Bin Li 等ICLR 2021 · 被引用 7,353 次
- Occlusion-Aware Networks for 3D Human Pose Estimation in VideoYu Cheng, Bo Yang, Bo Wang, Wending Yan 等ICCV 2019 · 被引用 223 次
- FocalClick: Towards Practical Interactive Image SegmentationXi Chen, Zhiyan Zhao, Yilei Zhang, Manni Duan 等CVPR 2022 · 被引用 153 次
- Amodal Segmentation Based on Visible Region Segmentation and Shape PriorYuting Xiao, Yanyu Xu, Ziming Zhong, Weixin Luo 等AAAI 2021 · 被引用 76 次
相关 Paper
- Amodal Instance Segmentation via Prior-Guided ExpansionJunjie Chen, Li Niu, Jianfu Zhang, Jianlou Si 等AAAI 2023 · 被引用 25 次
- Amodal Segmentation through Out-of-Task and Out-of-Distribution Generalization with a Bayesian ModelYihong Sun, Adam Kortylewski, Alan L. YuilleCVPR 2022 · 被引用 26 次
- Interactive Image Segmentation With First Click AttentionZheng Lin, Zhao Zhang, Lin-Zhuo Chen, Ming-Ming Cheng 等CVPR 2020
- BLADE: Box-Level Supervised Amodal Segmentation through Directed ExpansionZhaochen Liu, Zhixuan Li, Tingting JiangAAAI 2024 · 被引用 12 次
- Using Diffusion Priors for Video Amodal SegmentationKaihua Chen, Deva Ramanan, Tarasha KhuranaCVPR 2025
