Meta Parsing Networks: Towards Generalized Few-shot Scene Parsing with Adaptive Metric Learning
Peike Li, Yunchao Wei, Yi Yang
Abstract
Recent progress in few-shot segmentation usually aims at performing novel object segmentation using a few annotated examples as guidance. In this work, we advance this few-shot segmentation paradigm towards a more challenging yet general scenario, i.e., Generalized Few-shot Scene Parsing (GFSP). In this task, we take a fully annotated image as guidance to segment all pixels in a query image. Our mission is to study a generalizable and robust segmentation network from the meta-learning perspective so that both seen and unseen categories can be correctly recognized. Different from previous practices, this task performs segmentation on a joint label space consisting of both previously seen and novel categories. Moreover, pixels from these multiple categories need to be simultaneously taken into account, which is actually not well explored before. Accordingly, we present Meta Parsing Networks (MPNet) to better exploit the guidance information in the support set. Our MPNet contains two basic modules, i.e., the Adaptive Deep Metric Learning (ADML) module and the Contrastive Inter-class Distraction (CID) module. Specially, the ADML takes the annotated pixels from the support image as the guidance and adaptively produces high-quality prototypes for learning a deep comparison metric. In addition, MPNet further introduces the CID module learning to enlarge the feature discrepancy of different categories in the embedding space, leading the MPNet to generate more discriminative feature embeddings. We conduct experiments on two newly constructed benchmarks, i.e., GFSP-Cityscapes and GFSP-Pascal-Context. Extensive ablation studies well demonstrate the effectiveness and generalization ability of our MPNet.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 21ebf53c-0850-4322-8d43-167a6a3ace87Cited by top-tier papers4
- University-1652: A Multi-view Multi-source Benchmark for Drone-based Geo-localizationZhedong Zheng, Yunchao Wei, Yi YangACM MM 2020 · 390 citations
- Consistent Structural Relation Learning for Zero-Shot SegmentationPeike Li, Yunchao Wei, Yi YangNeurIPS 2020 · 88 citations
- Towards Cross-Granularity Few-Shot Learning: Coarse-to-Fine Pseudo-Labeling with Visual-Semantic Meta-EmbeddingJinhai Yang, Hua Yang, Lin ChenACM MM 2021 · 16 citations
- Few-Shot Multi-Agent PerceptionChenyou Fan, Junjie Hu, Jianwei HuangACM MM 2021 · 6 citations
Related papers
- ABPNet: Adaptive Background Modeling for Generalized Few Shot SegmentationKaiqi Dong, Wei Yang, Zhenbo Xu, Liusheng Huang et al.ACM MM 2021 · 12 citations
- Generalized Few-shot Semantic SegmentationZhuotao Tian, Xin Lai, Li Jiang, Shu Liu et al.CVPR 2022 · 103 citations
- Adaptive FSS: A Novel Few-Shot Segmentation Framework via Prototype EnhancementJing Wang, Jiangyun Li, Chen Chen, Yisi Zhang et al.AAAI 2024 · 24 citations
- PANet: Few-Shot Image Semantic Segmentation With Prototype AlignmentKaixin Wang, Jun Hao Liew, Yingtian Zou, Daquan Zhou et al.ICCV 2019 · 1,404 citations
- Learning What Not to Segment: A New Perspective on Few-Shot SegmentationChunbo Lang, Gong Cheng, Binfei Tu, Junwei HanCVPR 2022 · 289 citations
