Few-Shot Segmentation via Cycle-Consistent Transformer
Gengwei Zhang, Guoliang Kang, Yi Yang, Yunchao Wei
摘要
Few-shot segmentation aims to train a segmentation model that can fast adapt to novel classes with few exemplars. The conventional training paradigm is to learn to make predictions on query images conditioned on the features from support images. Previous methods only utilized the semantic-level prototypes of support images as conditional information. These methods cannot utilize all pixel-wise support information for the query predictions, which is however critical for the segmentation task. In this paper, we focus on utilizing pixel-wise relationships between support and query images to facilitate the few-shot segmentation task. We design a novel Cycle-Consistent TRansformer (CyCTR) module to aggregate pixel-wise support features into query ones. CyCTR performs cross-attention between features from different images, i.e. support and query images. We observe that there may exist unexpected irrelevant pixel-level support features. Directly performing cross-attention may aggregate these features from support to query and bias the query features. Thus, we propose using a novel cycle-consistent attention mechanism to filter out possible harmful support features and encourage query features to attend to the most informative pixels from support images. Experiments on all few-shot segmentation benchmarks demonstrate that our proposed CyCTR leads to remarkable improvement compared to previous state-of-the-art methods. Specifically, on Pascal- and COCO- datasets, we achieve 67.5% and 45.6% mIoU for 5-shot segmentation, outperforming previous state-of-the-art methods by 5.6% and 7.1% respectively.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper54
- Visual Prompting via Image InpaintingAmir Bar, Yossi Gandelsman, Trevor Darrell, Amir Globerson 等NeurIPS 2022 · 被引用 340 次
- SLCA: Slow Learner with Classifier Alignment for Continual Learning on a Pre-trained ModelGengwei Zhang, Liyuan Wang, Guoliang Kang, Ling Chen 等ICCV 2023 · 被引用 196 次
- Intermediate Prototype Mining Transformer for Few-Shot Semantic SegmentationYuanwei Liu, Nian Liu, Xiwen Yao, Junwei HanNeurIPS 2022 · 被引用 107 次
- Feature-Proxy Transformer for Few-Shot SegmentationJian-Wei Zhang, Yifan Sun, Yi Yang, Wei ChenNeurIPS 2022 · 被引用 105 次
- Singular Value Fine-tuning: Few-shot Segmentation requires Few-parameters Fine-tuningYanpeng Sun, Qiang Chen, Xiangyu He, Jian Wang 等NeurIPS 2022 · 被引用 97 次
它引用的顶会 Paper12
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Training data-efficient image transformers & distillation through attentionHugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa 等ICML 2021 · 被引用 8,974 次
- Deformable DETR: Deformable Transformers for End-to-End Object DetectionXizhou Zhu, Weijie Su, Lewei Lu, Bin Li 等ICLR 2021 · 被引用 7,353 次
- CCNet: Criss-Cross Attention for Semantic SegmentationZilong Huang, Xinggang Wang, Lichao Huang, Chang Huang 等ICCV 2019 · 被引用 2,972 次
- PANet: Few-Shot Image Semantic Segmentation With Prototype AlignmentKaixin Wang, Jun Hao Liew, Yingtian Zou, Daquan Zhou 等ICCV 2019 · 被引用 1,404 次
相关 Paper
- Dynamic Transformer for Few-shot Instance SegmentationHaochen Wang, Jie Liu, Yongtuo Liu, Subhransu Maji 等ACM MM 2022 · 被引用 10 次
- Mask Matching Transformer for Few-Shot SegmentationSiyu Jiao, Gengwei Zhang, Shant Navasardyan, Ling Chen 等NeurIPS 2022 · 被引用 54 次
- Unlocking the Potential of Pre-Trained Vision Transformers for Few-Shot Semantic Segmentation through Relationship DescriptorsZiqin Zhou, Hai-Ming Xu, Yangyang Shu, Lingqiao LiuCVPR 2024 · 被引用 7 次
- Few-Shot Semantic Segmentation with Cyclic Memory NetworkGuo-Sen Xie, Huan Xiong, Jie Liu, Yazhou Yao 等ICCV 2021 · 被引用 70 次
- CRNet: Cross-Reference Networks for Few-Shot SegmentationWeide Liu, Chi Zhang, Guosheng Lin, Fayao LiuCVPR 2020
