Semi-Supervised Video Salient Object Detection Based on Uncertainty-Guided Pseudo Labels
Yongri Piao, Chenyang Lu, Miao Zhang, Huchuan Lu
Abstract
Semi-Supervised Video Salient Object Detection (SS-VSOD) is challenging because of the lack of temporal information caused by sparse annotations in video sequences. Most works address this problem by generating pseudo labels for unlabeled data. However, error-prone pseudo labels negatively affect the VOSD model. Therefore, a deeper insight into pseudo labels should be developed. In this work, we aim to explore 1) how to utilize the incorrect predictions in pseudo labels to guide the network to generate more robust pseudo labels and 2) how to further screen out the noise that still exists in the improved pseudo labels. To this end, we propose an Uncertainty-Guided Pseudo Label Generator (UGPLG), which makes full use of inter-frame information to ensure the temporal consistency of the pseudo-labels and improves the robustness of the pseudo labels by strengthening the learning of difficult scenarios. Furthermore, we also introduce adversarial learning to address the noise problems in pseudo labels, guaranteeing the positive guidance of pseudo labels during model training. Experimental results demonstrate that our methods outperform existing semi-supervised method and partial fully-supervised methods across five public benchmarks of DAVIS, FBMS, MCL, ViSal, and SegTrack-V2.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e5efd055-73bf-4705-b5fe-e4171524659dCited by top-tier papers2
- Discover and Align Taxonomic Context Priors for Open-world Semi-Supervised LearningYu Wang, Zhun Zhong, Pengchong Qiao, Xuxin Cheng et al.NeurIPS 2023 · 25 citations
- Samba: A Unified Mamba-based Framework for General Salient Object DetectionJiahao He, Keren Fu, Xiaohong Liu, Qijun ZhaoCVPR 2025
Builds on15
- Depth-Induced Multi-Scale Recurrent Attention Network for Saliency DetectionYongri Piao, Wei Ji, Jingjing Li, Miao Zhang et al.ICCV 2019 · 450 citations
- Dual Student: Breaking the Limits of the Teacher in Semi-Supervised LearningZhanghan Ke, Daoye Wang, Qiong Yan, Jimmy S. J. Ren et al.ICCV 2019 · 259 citations
- Motion Guided Attention for Video Salient Object DetectionHaofeng Li, Guanqi Chen, Guanbin Li, Yizhou YuICCV 2019 · 200 citations
- Pyramid Constrained Self-Attention Network for Fast Video Salient Object DetectionYuchao Gu, Lijuan Wang, Ziqin Wang, Yun Liu et al.AAAI 2020 · 184 citations
- Full-Duplex Strategy for Video Object SegmentationGe-Peng Ji, Keren Fu, Zhe Wu, Deng-Ping Fan et al.ICCV 2021 · 173 citations
Related papers
- Semi-Supervised Video Salient Object Detection Using Pseudo-LabelsPengxiang Yan, Guanbin Li, Yuan Xie, Zhen Li et al.ICCV 2019 · 134 citations
- Learning from Noisy Pseudo Labels for Semi-Supervised Temporal Action LocalizationKun Xia, Le Wang, Sanping Zhou, Gang Hua et al.ICCV 2023 · 16 citations
- Towards Robust Video Object Segmentation with Adaptive Object CalibrationXiaohao Xu, Jinglu Wang, Xiang Ming, Yan LuACM MM 2022 · 21 citations
- In Defense of Pseudo-Labeling: An Uncertainty-Aware Pseudo-label Selection Framework for Semi-Supervised LearningMamshad Nayeem Rizve, Kevin Duarte, Yogesh S. Rawat, Mubarak ShahICLR 2021 · 630 citations
- Data-Uncertainty Guided Multi-Phase Learning for Semi-Supervised Object DetectionZhenyu Wang, Yali Li, Ye Guo, Lu Fang et al.CVPR 2021
