Rethinking Transformer-Based Blind-Spot Network for Self-Supervised Image Denoising
Junyi Li, Zhilu Zhang, Wangmeng Zuo
Abstract
Blind-spot networks (BSN) have been prevalent neural architectures in self-supervised image denoising (SSID). However, most existing BSNs are conducted with convolution layers. Although transformers have shown the potential to overcome the limitations of convolutions in many image restoration tasks, the attention mechanisms may violate the blind-spot requirement, thereby restricting their applicability in BSN. To this end, we propose to analyze and redesign the channel and spatial attentions to meet the blind-spot requirement. Specifically, channel self-attention may leak the blind-spot information in multi-scale architectures, since the downsampling shuffles the spatial feature into channel dimensions. To alleviate this problem, we divide the channel into several groups and perform channel attention separately. For spatial self-attention, we apply an elaborate mask to the attention matrix to restrict and mimic the receptive field of dilated convolution. Based on the redesigned channel and window attentions, we build a Transformer-based Blind-Spot Network (TBSN), which shows strong local fitting and global perspective abilities. Furthermore, we introduce a knowledge distillation strategy that distills TBSN into smaller denoisers to improve computational efficiency while maintaining performance. Extensive experiments on real-world image denoising datasets show that TBSN largely extends the receptive field and exhibits favorable performance against state-of-the-art SSID methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 43def5bc-1abc-4b25-8cc7-a2c69c4f614dCited by top-tier papers7
- APR-RD: Complemental Two Steps for Self-Supervised Real Image DenoisingHyunjun Kim, Nam Ik ChoAAAI 2025 · 6 citations
- Next-Scale Prediction: A Self-Supervised Approach for Real-World Image DenoisingYiwen Shan, Haiyu Zhao, Peng Hu, Xi Peng et al.CVPR 2026 · 2 citations
- Degradation-Aware Metric Prompting for Hyperspectral Image RestorationBinfeng Wang, Di Wang, Haonan Guo, Ying Fu et al.ICML 2026 · 2 citations
- TM-BSN: Triangular-Masked Blind-Spot Network for Real-World Self-Supervised Image DenoisingJunyoung Park, Youngjin Oh, Nam Ik ChoCVPR 2026 · 1 citation
- Zero-Shot Blind-Spot Image Denoising via Cross-Scale Non-Local Pixel RefillingQilong Guo, Tianjing Zhang, Zhiyuan Ma, Hui JiNeurIPS 2025
Builds on27
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Restormer: Efficient Transformer for High-Resolution Image RestorationSyed Waqas Zamir, Aditya Arora, Salman Khan, Munawar Hayat et al.CVPR 2022 · 3,348 citations
- Uformer: A General U-Shaped Transformer for Image RestorationZhendong Wang, Xiaodong Cun, Jianmin Bao, Wengang Zhou et al.CVPR 2022 · 1,970 citations
- AP-BSN: Self-Supervised Denoising for Real-World Images via Asymmetric PD and Blind-Spot NetworkWooseok Lee, Sanghyun Son, Kyoung Mu LeeCVPR 2022 · 148 citations
Related papers
- Exploring Efficient Asymmetric Blind-Spots for Self-Supervised Denoising in Real-World ScenariosShiyan Chen, Jiyuan Zhang, Zhaofei Yu, Tiejun HuangCVPR 2024
- Spatiotemporal Blind-Spot Network with Calibrated Flow Alignment for Self-Supervised Video DenoisingZikang Chen, Tao Jiang, Xiaowan Hu, Wang Zhang et al.AAAI 2025 · 3 citations
- Pseudo-Siamese Blind-spot Transformers for Self-Supervised Real-World DenoisingYuhui Quan, Tianxiang Zheng, Hui JiNeurIPS 2024 · 5 citations
- LG-BPN: Local and Global Blind-Patch Network for Self-Supervised Real-World DenoisingZichun Wang, Ying Fu, Ji Liu, Yulun ZhangCVPR 2023
- Self-supervised Image Denoising with Downsampled Invariance Loss and Conditional Blind-Spot NetworkYeong Il Jang, Keuntek Lee, Gu Yong Park, Seyun Kim et al.ICCV 2023 · 29 citations
