Triple-Cooperative Video Shadow Detection
Zhihao Chen, Liang Wan, Lei Zhu, Jia Shen, Huazhu Fu, Wennan Liu, Jing Qin
Abstract
Shadow detection in a single image has received significant research interests in recent years. However, much fewer works have been explored in shadow detection over dynamic scenes. The bottleneck is the lack of a wellestablished dataset with high-quality annotations for video shadow detection. In this work, we collect a new video shadow detection dataset (ViSha), which contains 120 videos with 11, 685 frames, covering 60 object categories, varying lengths, and different motion/lighting conditions. All the frames are annotated with a high-quality pixel-level shadow mask. To the best of our knowledge, this is the first learning-oriented dataset for video shadow detection. Furthermore, we develop a new baseline model, named triplecooperative video shadow detection network (TVSD-Net). It utilizes triple parallel networks in a cooperative manner to learn discriminative representations at intra-video and inter-video levels. Within the network, a dual gated co-attention module is proposed to constrain features from neighboring frames in the same video, while an auxiliary similarity loss is introduced to mine semantic information between different videos. Finally, we conduct a comprehensive study on ViSha, evaluating 12 state-of-the-art models (including single image shadow detectors, video object segmentation, and saliency detection methods). Experiments demonstrate that our model outperforms SOTA competitors.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4c2e7724-8ff9-45e7-8023-772479f53153Cited by top-tier papers13
- VIL-100: A New Dataset and A Baseline Model for Video Instance Lane DetectionYujun Zhang, Lei Zhu, Wei Feng, Huazhu Fu et al.ICCV 2021 · 67 citations
- Video Shadow Detection via Spatio-Temporal Interpolation Consistency TrainingXiao Lu, Yihong Cao, Sheng Liu, Chengjiang Long et al.CVPR 2022 · 24 citations
- Timeline and Boundary Guided Diffusion Network for Video Shadow DetectionHaipeng Zhou, Hongqiu Wang, Tian Ye, Zhaohu Xing et al.ACM MM 2024 · 18 citations
- SDDNet: Style-guided Dual-layer Disentanglement Network for Shadow DetectionRunmin Cong, Yuchen Guan, Jinpeng Chen, Wei Zhang et al.ACM MM 2023 · 15 citations
- Multi-View Dynamic Reflection Prior for Video Glass Surface DetectionFang Liu, Yuhao Liu, Jiaying Lin, Ke Xu et al.AAAI 2024 · 12 citations
Builds on3
- Video Object Segmentation Using Space-Time Memory NetworksSeoung Wug Oh, Joon-Young Lee, Ning Xu, Seon Joo KimICCV 2019 · 845 citations
- Motion Guided Attention for Video Salient Object DetectionHaofeng Li, Guanqi Chen, Guanbin Li, Yizhou YuICCV 2019 · 200 citations
- A Multi-Task Mean Teacher for Semi-Supervised Shadow DetectionZhihao Chen, Lei Zhu, Liang Wan, Song Wang et al.CVPR 2020
Related papers
- Semi-supervised Video Shadow Detection via Image-assisted Pseudo-label GenerationZipei Chen, Xiao Lu, Ling Zhang, Chunxia XiaoACM MM 2022 · 9 citations
- Language-Driven Interactive Shadow DetectionHongqiu Wang, Wei Wang, Haipeng Zhou, Huihui Xu et al.ACM MM 2024 · 7 citations
- SCOTCH and SODA: A Transformer Video Shadow Detection FrameworkLihao Liu, Jean Prost, Lei Zhu, Nicolas Papadakis et al.CVPR 2023
- DTTNet: Improving Video Shadow Detection via Dark-Aware Guidance and Tokenized Temporal ModelingZhicheng Li, Kunyang Sun, Rui Yao, Hancheng Zhu et al.AAAI 2026
- OmniSR: Shadow Removal Under Direct and Indirect LightingJiamin Xu, Zelong Li, Yuxin Zheng, Chenyu Huang et al.AAAI 2025 · 18 citations
