Suppressing the Heterogeneity: A Strong Feature Extractor for Few-shot Segmentation
Zhengdong Hu, Yifan Sun, Yi Yang
摘要
This paper tackles the Few-shot Semantic Segmentation (FSS) task with focus on learning the feature extractor. Somehow the feature extractor has been overlooked by recent state-of-the-art methods, which directly use a deep model pretrained on ImageNet for feature extraction (without further fine-tuning). Under this background, we think the FSS feature extractor deserves exploration and observe the heterogeneity (i.e., the intra-class diversity in the raw images) as a critical challenge hindering the intra-class feature compactness. The heterogeneity has three levels from coarse to fine: 1) Sample-level: the inevitable distribution gap between the support and query images makes them heterogeneous from each other. 2) Region-level: the background in FSS actually contains multiple regions with different semantics. 3) Patch-level: some neighboring patches belonging to a same class may appear quite different from each other. Motivated by these observations, we propose a feature extractor with Multi-level Heterogeneity Suppressing (MuHS). MuHS leverages the attention mechanism in transformer backbone to effectively suppress all these three-level heterogeneity. Concretely, MuHS reinforces the attention / interaction between different samples (query and support), different regions and neighboring patches by constructing cross-sample attention, cross-region interaction and a novel masked image segmentation (inspired by the recent masked image modeling), respectively. We empirically show that 1) MuHS brings consistent improvement for various FSS heads and 2) using a simple linear classification head, MuHS sets new states of the art on multiple FSS datasets, validating the importance of FSS feature learning.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper10
- Knowledge-Enhanced Dual-Stream Zero-Shot Composed Image RetrievalYucheng Suo, Fan Ma, Linchao Zhu, Yi YangCVPR 2024 · 被引用 20 次
- Learning Solution-Aware Transformers for Efficiently Solving Quadratic Assignment ProblemZhentao Tan, Yadong MuICML 2024 · 被引用 5 次
- Object-Level Correlation for Few-Shot SegmentationChunlin Wen, Yu Zhang, Jie Fan, Hongyuan Zhu 等ICCV 2025 · 被引用 5 次
- Enhancing Generalized Few-Shot Semantic Segmentation via Effective Knowledge TransferXinyue Chen, Miaojing Shi, Zijian Zhou, Lianghua He 等AAAI 2025 · 被引用 3 次
- Unified Mask Embedding and Correspondence Learning for Self-Supervised Video SegmentationLiulei Li, Wenguan Wang, Tianfei Zhou, Jianwu Li 等CVPR 2023
相关 Paper
- Feature-Proxy Transformer for Few-Shot SegmentationJian-Wei Zhang, Yifan Sun, Yi Yang, Wei ChenNeurIPS 2022 · 被引用 105 次
- Hierarchical Dense Correlation Distillation for Few-Shot SegmentationBohao Peng, Zhuotao Tian, Xiaoyang Wu, Chengyao Wang 等CVPR 2023
- Simpler is Better: Few-shot Semantic Segmentation with Classifier Weight TransformerZhihe Lu, Sen He, Xiatian Zhu, Li Zhang 等ICCV 2021 · 被引用 232 次
- Integrative Few-Shot Learning for Classification and SegmentationDahyun Kang, Minsu ChoCVPR 2022 · 被引用 76 次
- Semantic Prompt for Few-Shot Image RecognitionCVPR 2023
