Intermediate Prototype Mining Transformer for Few-Shot Semantic Segmentation
Yuanwei Liu, Nian Liu, Xiwen Yao, Junwei Han
摘要
Few-shot semantic segmentation aims to segment the target objects in query under the condition of a few annotated support images. Most previous works strive to mine more effective category information from the support to match with the corresponding objects in query. However, they all ignored the category information gap between query and support images. If the objects in them show large intraclass diversity, forcibly migrating the category information from the support to the query is ineffective. To solve this problem, we are the first to introduce an intermediate prototype for mining both deterministic category information from the support and adaptive category knowledge from the query. Specifically, we design an Intermediate Prototype Mining Transformer (IPMT) to learn the prototype in an iterative way. In each IPMT layer, we propagate the object information in both support and query features to the prototype and then use it to activate the query feature map. By conducting this process iteratively, both the intermediate prototype and the query feature can be progressively improved. At last, the final query feature is used to yield precise segmentation prediction. Extensive experiments on both PASCAL-5 i and COCO-20 i datasets clearly verify the effectiveness of our IPMT and show that it outperforms previous state-of-the-art methods by a large margin. Code is available at https://github.com/LIUYUANWEI98/IPMT
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper24
- Self-Calibrated Cross Attention Network for Few-Shot SegmentationQianxiong Xu, Wenting Zhao, Guosheng Lin, Cheng LongICCV 2023 · 被引用 76 次
- Bridge the Points: Graph-based Few-shot Segment Anything SemanticallyAnqi Zhang, Guangyu Gao, Jianbo Jiao, Chi Harold Liu 等NeurIPS 2024 · 被引用 56 次
- LLaFS: When Large Language Models Meet Few-Shot SegmentationLanyun Zhu, Tianrun Chen, Deyi Ji, Jieping Ye 等CVPR 2024 · 被引用 39 次
- Relevant Intrinsic Feature Enhancement Network for Few-Shot Semantic SegmentationXiaoyi Bao, Jie Qin, Siyang Sun, Xingang Wang 等AAAI 2024 · 被引用 34 次
- Focus on Query: Adversarial Mining Transformer for Few-Shot SegmentationYuan Wang, Naisong Luo, Tianzhu ZhangNeurIPS 2023 · 被引用 29 次
它引用的顶会 Paper22
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Deformable DETR: Deformable Transformers for End-to-End Object DetectionXizhou Zhu, Weijie Su, Lewei Lu, Bin Li 等ICLR 2021 · 被引用 7,353 次
- Pyramid Vision Transformer: A Versatile Backbone for Dense Prediction without ConvolutionsWenhai Wang, Enze Xie, Xiang Li, Deng-Ping Fan 等ICCV 2021 · 被引用 4,909 次
- PANet: Few-Shot Image Semantic Segmentation With Prototype AlignmentKaixin Wang, Jun Hao Liew, Yingtian Zou, Daquan Zhou 等ICCV 2019 · 被引用 1,404 次
相关 Paper
- Adaptive FSS: A Novel Few-Shot Segmentation Framework via Prototype EnhancementJing Wang, Jiangyun Li, Chen Chen, Yisi Zhang 等AAAI 2024 · 被引用 24 次
- Bidirectional Reciprocative Information Communication for Few-Shot Semantic SegmentationYuanwei Liu, Junwei Han, Xiwen Yao, Salman Khan 等ICML 2024 · 被引用 6 次
- Multi-grained Temporal Prototype Learning for Few-shot Video Object SegmentationNian Liu, Kepan Nan, Wangbo Zhao, Yuanwei Liu 等ICCV 2023 · 被引用 12 次
- Mask Matching Transformer for Few-Shot SegmentationSiyu Jiao, Gengwei Zhang, Shant Navasardyan, Ling Chen 等NeurIPS 2022 · 被引用 54 次
- Learning from the Target: Dual Prototype Network for Few Shot Semantic SegmentationBinjie Mao, Xinbang Zhang, Lingfeng Wang, Qian Zhang 等AAAI 2022 · 被引用 23 次
