Multi-grained Temporal Prototype Learning for Few-shot Video Object Segmentation
Nian Liu, Kepan Nan, Wangbo Zhao, Yuanwei Liu, Xiwen Yao, Salman Khan, Hisham Cholakkal, Rao Muhammad Anwer, Junwei Han, Fahad Shahbaz Khan
摘要
Few-Shot Video Object Segmentation (FSVOS) aims to segment objects in a query video with the same category defined by a few annotated support images. However, this task was seldom explored. In this work, based on IPMT, a state-of-the-art few-shot image segmentation method that combines external support guidance information with adaptive query guidance cues, we propose to leverage multi-grained temporal guidance information for handling the temporal correlation nature of video data. We decompose the query video information into a clip prototype and a memory prototype for capturing local and long-term internal temporal guidance, respectively. Frame prototypes are further used for each frame independently to handle fine-grained adaptive guidance and enable bidirectional clip-frame prototype communication. To reduce the influence of noisy memory, we propose to leverage the structural similarity relation among different predicted regions and the support for selecting reliable memory frames. Furthermore, a new segmentation loss is also proposed to enhance the category discriminability of the learned prototypes. Experimental results demonstrate that our proposed video IPMT model significantly outperforms previous models on two benchmark datasets. Code is available at https://github.com/nankepan/VIPMT.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Rethinking Prior Information Generation with CLIP for Few-Shot SegmentationJin Wang, Bingfeng Zhang, Jian Pang, Honglong Chen 等CVPR 2024 · 被引用 27 次
- Bidirectional Reciprocative Information Communication for Few-Shot Semantic SegmentationYuanwei Liu, Junwei Han, Xiwen Yao, Salman Khan 等ICML 2024 · 被引用 6 次
- MOVE: Motion-Guided Few-Shot Video Object SegmentationKaining Ying, Hengrui Hu, Henghui DingICCV 2025 · 被引用 3 次
- Beyond Pixel and Object: Part Feature as Reference for Few-Shot Video Object SegmentationNaisong Luo, Guoxin Xiong, Tianzhu ZhangAAAI 2025
- GP-NeRF: Generalized Perception NeRF for Context-Aware 3D Scene UnderstandingHao Li, Dingwen Zhang, Yalun Dai, Nian Liu 等CVPR 2024
它引用的顶会 Paper12
- Per-Pixel Classification is Not All You Need for Semantic SegmentationBowen Cheng, Alexander G. Schwing, Alexander KirillovNeurIPS 2021 · 被引用 2,196 次
- Video Object Segmentation Using Space-Time Memory NetworksSeoung Wug Oh, Joon-Young Lee, Ning Xu, Seon Joo KimICCV 2019 · 被引用 845 次
- Video Instance SegmentationLinjie Yang, Yuchen Fan, Ning XuICCV 2019 · 被引用 615 次
- Rethinking Space-Time Networks with Improved Memory Coverage for Efficient Video Object SegmentationHo Kei Cheng, Yu-Wing Tai, Chi-Keung TangNeurIPS 2021 · 被引用 403 次
- Feature Weighting and Boosting for Few-Shot SegmentationKhoi Nguyen, Sinisa TodorovicICCV 2019 · 被引用 402 次
相关 Paper
- Intermediate Prototype Mining Transformer for Few-Shot Semantic SegmentationYuanwei Liu, Nian Liu, Xiwen Yao, Junwei HanNeurIPS 2022 · 被引用 107 次
- Multi-modal Prototype Guided Few-shot Object DetectionChenbo Zhang, Bing Huangfu, Hongxu Ma, Jihong Guan 等ACM MM 2025 · 被引用 3 次
- Delving Deep Into Many-to-Many Attention for Few-Shot Video Object SegmentationHaoxin Chen, Hanjie Wu, Nanxuan Zhao, Sucheng Ren 等CVPR 2021
- MPL: Match-guided Prototype Learning for Few-shot Action RecognitionFeng Yang, Jie Zhao, Fulin Luo, Anyong Qin 等CVPR 2026
- Few-Shot Semantic Segmentation with Cyclic Memory NetworkGuo-Sen Xie, Huan Xiong, Jie Liu, Yazhou Yao 等ICCV 2021 · 被引用 70 次
