Stochastic Window Transformer for Image Restoration
Jie Xiao, Xueyang Fu, Feng Wu, Zheng-Jun Zha
Abstract
Thanks to the powerful representation capabilities, transformers have made impressive progress in image restoration. However, existing transformers-based methods do not carefully consider the particularities of image restoration. In general, image restoration requires that an ideal approach should be translation-invariant to the degradation, i.e., the undesirable degradation should be removed irrespective of its position within the image. Furthermore, the local relationships also play a vital role, which should be faithfully exploited for recovering clean images. Nevertheless, most transformers either adopt local attention with the fixed local window strategy or global attention, which unfortunately breaks the translation invariance and causes huge loss of local relationships. To address these issues, we propose an elegant stochastic window strategy for transformers. Specifically, we first introduce the window partition with stochastic shift to replace the original fixed window partition for training. Then, we design a new layer expectation propagation algorithm to efficiently approximate the expectation of the induced stochastic transformer for testing. Our stochastic window transformer not only enjoys powerful representation but also maintains the desired property of translation invariance and locality. Experiments validate the stochastic window strategy consistently improves performance on various image restoration tasks (derain-ing, denoising and deblurring) by significant margins. The code is available at https://github.com/jiexiaou/Stoformer .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 62ab7f63-2d60-4b4a-87e9-b12738985b19Cited by top-tier papers7
- Image Restoration with Mean-Reverting Stochastic Differential EquationsZiwei Luo, Fredrik K. Gustafsson, Zheng Zhao, Jens Sjölund et al.ICML 2023 · 293 citations
- ESSAformer: Efficient Transformer for Hyperspectral Image Super-resolutionMingjin Zhang, Chi Zhang, Qiming Zhang, Jie Guo et al.ICCV 2023 · 73 citations
- Random Shuffle Transformer for Image RestorationJie Xiao, Xueyang Fu, Man Zhou, Hongjian Liu et al.ICML 2023 · 38 citations
- Motion-adaptive Transformer for Event-based Image DeblurringSenyan Xu, Zhijing Sun, Mingchen Zhong, Chengzhi Cao et al.AAAI 2025 · 17 citations
- EventMamba: Enhancing Spatio-Temporal Locality with State Space Models for Event-Based Video ReconstructionChengjie Ge, Xueyang Fu, Peng He, Kunyu Wang et al.AAAI 2025 · 6 citations
Builds on24
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Training data-efficient image transformers & distillation through attentionHugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa et al.ICML 2021 · 8,974 citations
- Pyramid Vision Transformer: A Versatile Backbone for Dense Prediction without ConvolutionsWenhai Wang, Enze Xie, Xiang Li, Deng-Ping Fan et al.ICCV 2021 · 4,909 citations
- Restormer: Efficient Transformer for High-Resolution Image RestorationSyed Waqas Zamir, Aditya Arora, Salman Khan, Munawar Hayat et al.CVPR 2022 · 3,348 citations
Related papers
- Uformer: A General U-Shaped Transformer for Image RestorationZhendong Wang, Xiaodong Cun, Jianmin Bao, Wengang Zhou et al.CVPR 2022 · 1,970 citations
- Enhancing Image Restoration Transformer via Adaptive Translation EquivarianceJiaKui Hu, Zhengjian Yao, Lujia Jin, Hangzhou He et al.ICCV 2025 · 6 citations
- KNN Local Attention for Image RestorationHunsang Lee, Hyesong Choi, Kwanghoon Sohn, Dongbo MinCVPR 2022 · 62 citations
- Cross Aggregation Transformer for Image RestorationZheng Chen, Yulun Zhang, Jinjin Gu, Yongbing Zhang et al.NeurIPS 2022 · 274 citations
- Bend the Basics: Degradation-Aware Deformable Tokenization for All-in-One Image RestorationZihao He, Yunfeng Wu, Xinchao Wang, Songhua LiuICML 2026
