Learning A Sparse Transformer Network for Effective Image Deraining
Xiang Chen, Hao Li, Mingqiang Li, Jinshan Pan
Abstract
Transformers-based methods have achieved significant performance in image deraining as they can model the non-local information which is vital for high-quality image reconstruction. In this paper, we find that most existing Transformers usually use all similarities of the tokens from the query-key pairs for the feature aggregation. However, if the tokens from the query are different from those of the key, the self-attention values estimated from these tokens also involve in feature aggregation, which accordingly interferes with the clear image restoration. To overcome this problem, we propose an effective DeRaining network, Sparse Transformer (DRSformer) that can adaptively keep the most useful self-attention values for feature aggregation so that the aggregated features better facilitate high-quality image reconstruction. Specifically, we develop a learnable top-k selection operator to adaptively retain the most crucial attention scores from the keys for each query for better feature aggregation. Simultaneously, as the naive feed-forward network in Transformers does not model the multi-scale information that is important for latent clear image restoration, we develop an effective mixed-scale feed-forward network to generate better features for image deraining. To learn an enriched set of hybrid features, which combines local context from CNN operators, we equip our model with mixture of experts feature compensator to present a cooperation refinement deraining scheme. Extensive experimental results on the commonly used benchmarks demonstrate that the proposed method achieves favorable performance against state-of-the-art approaches. The source code and trained models are available at https://github. com/cschenxiang/DRSformer .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e27500ff-a0ef-4252-a2de-b04758fbeeaeCited by top-tier papers78
- Adapt or Perish: Adaptive Sparse Transformer with Attentive Feature Refinement for Image RestorationShihao Zhou, Duosheng Chen, Jinshan Pan, Jinglei Shi et al.CVPR 2024 · 137 citations
- DreamClear: High-Capacity Real-World Image Restoration with Privacy-Safe Dataset CurationYuang Ai, Xiaoqiang Zhou, Huaibo Huang, Xiaotian Han et al.NeurIPS 2024 · 81 citations
- Sparse Sampling Transformer with Uncertainty-Driven Ranking for Unified Removal of Raindrops and Rain StreaksSixiang Chen, Tian Ye, Jinbin Bai, Erkang Chen et al.ICCV 2023 · 72 citations
- I2EBench: A Comprehensive Benchmark for Instruction-based Image EditingYiwei Ma, Jiayi Ji, Ke Ye, Weihuang Lin et al.NeurIPS 2024 · 67 citations
- Mutual Information-driven Triple Interaction Network for Efficient Image DehazingHao Shen, Zhong-Qiu Zhao, Yulun Zhang, Zhao ZhangACM MM 2023 · 59 citations
Builds on27
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Restormer: Efficient Transformer for High-Resolution Image RestorationSyed Waqas Zamir, Aditya Arora, Salman Khan, Munawar Hayat et al.CVPR 2022 · 3,348 citations
- Uformer: A General U-Shaped Transformer for Image RestorationZhendong Wang, Xiaodong Cun, Jianmin Bao, Wengang Zhou et al.CVPR 2022 · 1,970 citations
- Incorporating Convolution Designs into Visual TransformersKun Yuan, Shaopeng Guo, Ziwei Liu, Aojun Zhou et al.ICCV 2021 · 581 citations
Related papers
- Rethinking Multi-Scale Representations in Deep Deraining TransformerHongming Chen, Xiang Chen, Jiyang Lu, Yufeng LiAAAI 2024 · 46 citations
- Hybrid CNN-Transformer Feature Fusion for Single Image DerainingXiang Chen, Jinshan Pan, Jiyang Lu, Zhentao Fan et al.AAAI 2023 · 75 citations
- Cross Paradigm Representation and Alignment Transformer for Image DerainingShun Zou, Yi Zou, Juncheng Li, Guangwei Gao et al.ACM MM 2025 · 19 citations
- Magic ELF: Image Deraining Meets Association Learning and TransformerKui Jiang, Zhongyuan Wang, Chen Chen, Zheng Wang et al.ACM MM 2022 · 99 citations
- Bidirectional Multi-Scale Implicit Neural Representations for Image DerainingXiang Chen, Jinshan Pan, Jiangxin DongCVPR 2024
