Hybrid CNN-Transformer Feature Fusion for Single Image Deraining
Xiang Chen, Jinshan Pan, Jiyang Lu, Zhentao Fan, Hao Li
Abstract
Since rain streaks exhibit diverse geometric appearances and irregular overlapped phenomena, these complex characteristics challenge the design of an effective single image deraining model. To this end, rich local-global information representations are increasingly indispensable for better satisfying rain removal. In this paper, we propose a lightweight Hybrid CNN-Transformer Feature Fusion Network (dubbed as HCT-FFN) in a stage-by-stage progressive manner, which can harmonize these two architectures to help image restoration by leveraging their individual learning strengths. Specifically, we stack a sequence of the degradation-aware mixture of experts (DaMoE) modules in the CNN-based stage, where appropriate local experts adaptively enable the model to emphasize spatially-varying rain distribution features. As for the Transformer-based stage, a background-aware vision Transformer (BaViT) module is employed to complement spatially-long feature dependencies of images, so as to achieve global texture recovery while preserving the required structure. Considering the indeterminate knowledge discrepancy among CNN features and Transformer features, we introduce an interactive fusion branch at adjacent stages to further facilitate the reconstruction of high-quality deraining results. Extensive evaluations show the effectiveness and extensibility of our developed HCT-FFN. The source code is available at https://github.com/cschenxiang/HCT-FFN.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext dc74b07f-ddc3-454a-9408-0e75351610c4Cited by top-tier papers11
- Mutual Information-driven Triple Interaction Network for Efficient Image DehazingHao Shen, Zhong-Qiu Zhao, Yulun Zhang, Zhao ZhangACM MM 2023 · 59 citations
- Rethinking Multi-Scale Representations in Deep Deraining TransformerHongming Chen, Xiang Chen, Jiyang Lu, Yufeng LiAAAI 2024 · 46 citations
- Boosting Image De-Raining via Central-Surrounding Synergistic ConvolutionLong Peng, Yang Wang, Xin Di, Peizhe Xia et al.AAAI 2025 · 27 citations
- Cross Paradigm Representation and Alignment Transformer for Image DerainingShun Zou, Yi Zou, Juncheng Li, Guangwei Gao et al.ACM MM 2025 · 19 citations
- FouriDown: Factoring Down-Sampling into Shuffling and SuperposingQi Zhu, Man Zhou, Jie Huang, Naishan Zheng et al.NeurIPS 2023 · 12 citations
Builds on20
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Restormer: Efficient Transformer for High-Resolution Image RestorationSyed Waqas Zamir, Aditya Arora, Salman Khan, Munawar Hayat et al.CVPR 2022 · 3,348 citations
- CvT: Introducing Convolutions to Vision TransformersHaiping Wu, Bin Xiao, Noel Codella, Mengchen Liu et al.ICCV 2021 · 2,397 citations
- Uformer: A General U-Shaped Transformer for Image RestorationZhendong Wang, Xiaodong Cun, Jianmin Bao, Wengang Zhou et al.CVPR 2022 · 1,970 citations
Related papers
- Magic ELF: Image Deraining Meets Association Learning and TransformerKui Jiang, Zhongyuan Wang, Chen Chen, Zheng Wang et al.ACM MM 2022 · 99 citations
- CLG-INet: Coupled Local-Global Interactive Network for Image RestorationYuqi Jiang, Chune Zhang, Shuo Jin, Jiao Liu et al.ACM MM 2023 · 4 citations
- Learning A Sparse Transformer Network for Effective Image DerainingXiang Chen, Hao Li, Mingqiang Li, Jinshan PanCVPR 2023
- Bidirectional Multi-Scale Implicit Neural Representations for Image DerainingXiang Chen, Jinshan Pan, Jiangxin DongCVPR 2024
- Multi-Scale Progressive Fusion Network for Single Image DerainingKui Jiang, Zhongyuan Wang, Peng Yi, Chen Chen et al.CVPR 2020
