Spatiotemporal Blind-Spot Network with Calibrated Flow Alignment for Self-Supervised Video Denoising
Zikang Chen, Tao Jiang, Xiaowan Hu, Wang Zhang, Huaqiu Li, Haoqian Wang
Abstract
Self-supervised video denoising aims to remove noise from videos without relying on ground truth data, leveraging the video itself to recover clean frames. Existing methods often rely on simplistic feature stacking or apply optical flow without thorough analysis. This results in suboptimal utilization of both inter-frame and intra-frame information, and it also neglects the potential of optical flow alignment under self-supervised conditions, leading to biased and insufficient denoising outcomes. To this end, we first explore the practicality of optical flow in the self-supervised setting and introduce a SpatioTemporal Blind-spot Network (STBN) for global frame feature utilization. In the temporal domain, we utilize bidirectional blind-spot feature propagation through the proposed blind-spot alignment block to ensure accurate temporal alignment and effectively capture long-range dependencies. In the spatial domain, we introduce the spatial receptive field expansion module, which enhances the receptive field and improves global perception capabilities. Additionally, to reduce the sensitivity of optical flow estimation to noise, we propose an unsupervised optical flow distillation mechanism that refines fine-grained inter-frame interactions during optical flow alignment. Our method demonstrates superior performance across both synthetic and realworld video denoising datasets. The source code is publicly available at https://github.com/ZKCCZ/STBN .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 128010c4-39c7-4121-b7ef-099503a60203Cited by top-tier papers1
Ask how each one uses itBuilds on13
- BasicVSR++: Improving Video Super-Resolution with Enhanced Propagation and AlignmentKelvin C. K. Chan, Shangchen Zhou, Xiangyu Xu, Chen Change LoyCVPR 2022 · 522 citations
- Recurrent Video Restoration Transformer with Guided Deformable AttentionJingyun Liang, Yuchen Fan, Xiaoyu Xiang, Rakesh Ranjan et al.NeurIPS 2022 · 318 citations
- Patch Craft: Video Denoising by Deep Modeling and Patch MatchingGregory Vaksman, Michael Elad, Peyman MilanfarICCV 2021 · 79 citations
- Unsupervised Deep Video DenoisingDev Yashpal Sheth, Sreyas Mohan, Joshua L. Vincent, Ramon Manzorro et al.ICCV 2021 · 78 citations
- NightLab: A Dual-level Architecture with Hardness Detection for Segmentation at NightXueqing Deng, Peng Wang, Xiaochen Lian, Shawn D. NewsamCVPR 2022 · 51 citations
Related papers
- Recurrent Self-Supervised Video Denoising with Denser Receptive FieldZichun Wang, Yulun Zhang, Debing Zhang, Ying FuACM MM 2023 · 15 citations
- Rethinking Transformer-Based Blind-Spot Network for Self-Supervised Image DenoisingJunyi Li, Zhilu Zhang, Wangmeng ZuoAAAI 2025 · 31 citations
- Exploring Efficient Asymmetric Blind-Spots for Self-Supervised Denoising in Real-World ScenariosShiyan Chen, Jiyuan Zhang, Zhaofei Yu, Tiejun HuangCVPR 2024
- Self-supervised Image Denoising with Downsampled Invariance Loss and Conditional Blind-Spot NetworkYeong Il Jang, Keuntek Lee, Gu Yong Park, Seyun Kim et al.ICCV 2023 · 29 citations
- TM-BSN: Triangular-Masked Blind-Spot Network for Real-World Self-Supervised Image DenoisingJunyoung Park, Youngjin Oh, Nam Ik ChoCVPR 2026 · 1 citation
