Dynamic Context-Sensitive Filtering Network for Video Salient Object Detection
Miao Zhang, Jie Liu, Yifei Wang, Yongri Piao, Shunyu Yao, Wei Ji, Jingjing Li, Huchuan Lu, Zhongxuan Luo
Abstract
The ability to capture inter-frame dynamics has been critical to the development of video salient object detection (VSOD). While many works have achieved great success in this field, a deeper insight into its dynamic nature should be developed. In this work, we aim to answer the following questions: How can a model adjust itself to dynamic variations as well as perceive fine differences in the real-world environment; How are the temporal dynamics well introduced into spatial information over time? To this end, we propose a dynamic context-sensitive filtering network (DCFNet) equipped with a dynamic context-sensitive filtering module (DCFM) and an effective bidirectional dynamic fusion strategy. The proposed DCFM sheds new light on dynamic filter generation by extracting location-related affinities between consecutive frames. Our bidirectional dynamic fusion strategy encourages the interaction of spatial and temporal information in a dynamic manner. Experimental results demonstrate that our proposed method can achieve state-of-the-art performance on most VSOD datasets while ensuring a real-time speed of 28 fps. The source code is publicly available at https://github.com/OIPLab-DUT/DCFNet.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0c3008ea-b36a-4c0b-a3ef-d074faa9572dCited by top-tier papers19
- Pyramid Grafting Network for One-Stage High Resolution Saliency DetectionChenxi Xie, Changqun Xia, Mingcan Ma, Zhirui Zhao et al.CVPR 2022 · 112 citations
- Unsupervised Domain Adaptation for Nighttime Aerial TrackingJunjie Ye, Changhong Fu, Guangze Zheng, Danda Pani Paudel et al.CVPR 2022 · 109 citations
- VSCode: General Visual Salient and Camouflaged Object Detection with 2D Prompt LearningZiyang Luo, Nian Liu, Wangbo Zhao, Xuguang Yang et al.CVPR 2024 · 96 citations
- Joint Semantic Mining for Weakly Supervised RGB-D Salient Object DetectionJingjing Li, Wei Ji, Qi Bi, Cheng Yan et al.NeurIPS 2021 · 56 citations
- Promoting Saliency From Depth: Deep Unsupervised RGB-D Saliency DetectionWei Ji, Jingjing Li, Qi Bi, Chuan Guo et al.ICLR 2022 · 46 citations
Builds on11
- Depth-Induced Multi-Scale Recurrent Attention Network for Saliency DetectionYongri Piao, Wei Ji, Jingjing Li, Miao Zhang et al.ICCV 2019 · 450 citations
- Dynamic Multi-Scale Filters for Semantic SegmentationJunjun He, Zhongying Deng, Yu QiaoICCV 2019 · 287 citations
- Spatio-Temporal Filter Adaptive Network for Video DeblurringShangchen Zhou, Jiawei Zhang, Jinshan Pan, Wangmeng Zuo et al.ICCV 2019 · 225 citations
- Motion Guided Attention for Video Salient Object DetectionHaofeng Li, Guanqi Chen, Guanbin Li, Yizhou YuICCV 2019 · 200 citations
- Pyramid Constrained Self-Attention Network for Fast Video Salient Object DetectionYuchao Gu, Lijuan Wang, Ziqin Wang, Yun Liu et al.AAAI 2020 · 184 citations
Related papers
- TASED-Net: Temporally-Aggregating Spatial Encoder-Decoder Network for Video Saliency DetectionKyle Min, Jason J. CorsoICCV 2019 · 189 citations
- Full-Duplex Strategy for Video Object SegmentationGe-Peng Ji, Keren Fu, Zhe Wu, Deng-Ping Fan et al.ICCV 2021 · 173 citations
- Bringing Events into Video Deblurring with Non-consecutively Blurry FramesWei Shang, Dongwei Ren, Dongqing Zou, Jimmy S. Ren et al.ICCV 2021 · 85 citations
- A Motion-aware Spatio-temporal Graph for Video Salient Object RankingHao Chen, Yufei Zhu, Yongjian DengNeurIPS 2024 · 1 citation
- MFNet: Multi-filter Directive Network for Weakly Supervised Salient Object DetectionYongri Piao, Jian Wang, Miao Zhang, Huchuan LuICCV 2021 · 64 citations
