VPSentry: Semi-supervised Video Polyp Segmentation via Sentry-guided Long-term Prototype Fusion with Correlation Dynamic Propagation
Guilian Chen, Xiaoling Luo, Huisi Wu, Jing Qin
Abstract
Automated polyp segmentation in colonoscopy videos is an essential computer-aided technology for early detection and removal of polyps. However, most existing video polyp segmentation methods are designed with pixel-level temporal learning mechanisms, at the cost of time-consuming frame-wise annotations. In this paper, we present VPSentry, a novel semi-supervised segmentation model with a sentry mechanism. Our model integrates a prototype memory to store the long-term spatiotemporal cues of colonoscopy videos. Moreover, we devise adaptive prototypes to capture and generalize critical representations from individual frames, enabling long-term temporal fusion across labeled and unlabeled frames. In addition, we propose a correlation dynamic propagation module that propagates information from prototypes to features while simultaneously extracting dynamic features to perceive variations in polyp details between adjacent frames. Since colonoscopy scenes may change among consecutive frames, we further employ a sentry mechanism to assess the inter-frame continuity. This mechanism guides the prototype memory updating and the correlation dynamic propagation, further facilitating robust temporal propagation and dynamic detail perception for semi-supervised learning of long-term colonoscopy video sequences. Extensive experiments on the large-scale SUN-SEG dataset demonstrate that our model achieves optimal segmentation performance with real-time inference efficiency.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on8
- DDP: Diffusion Model for Dense Visual PredictionYuanfeng Ji, Zhe Chen, Enze Xie, Lanqing Hong et al.ICCV 2023 · 223 citations
- ACL-Net: Semi-supervised Polyp Segmentation via Affinity Contrastive LearningHuisi Wu, Wende Xie, Jingyin Lin, Xinrong GuoAAAI 2023 · 32 citations
- Cross-View Mutual Learning for Semi-Supervised Medical Image SegmentationSong Wu, Xiaoyu Wei, Xinyue Chen, Yazhou Ren et al.ACM MM 2024 · 16 citations
- Adaptive Learning of High-Value Regions for Semi-Supervised Medical Image SegmentationTao Lei, Ziyao Yang, Xingwu Wang, Yi Wang et al.ICCV 2025 · 5 citations
- Dual-calibrated Co-training Framework for Personalized Federated Semi-Supervised Medical Image SegmentationDelin Pan, Jiansong Fan, Jie Zhu, Llihua Li et al.AAAI 2025 · 5 citations
Related papers
- STDDNet: Harnessing Mamba for Video Polyp Segmentation via Spatial-aligned Temporal Modeling and Discriminative Dynamic Representation LearningGuilian Chen, Huisi Wu, Jing QinICCV 2025 · 4 citations
- An Embedding-Unleashing Video Polyp Segmentation Framework via Region Linking and Scale AlignmentZhixue Fang, Xinrong Guo, Jingyin Lin, Huisi Wu et al.AAAI 2024 · 8 citations
- HFSTI-Net: Hierarchical Frequency-spatial-temporal Interactions for Video Polyp SegmentationYuanqin He, Guilian Chen, Yuhua Zhang, Huisi Wu et al.ICLR 2026
- WavePolyp: Video Polyp Segmentation via Hierarchical Wavelet-Based Feature Aggregation and Inter-Frame Divergence PerceptionYuhua Zhang, Guilian Chen, Yuanqin He, Huisi Wu et al.ICLR 2026
- Collaborative and Adversarial Learning of Focused and Dispersive Representations for Semi-supervised Polyp SegmentationHuisi Wu, Guilian Chen, Zhenkun Wen, Jing QinICCV 2021 · 55 citations
