Semantics and Content Matter: Towards Multi-Prior Hierarchical Mamba for Image Deraining
Zhaocheng Yu, Kui Jiang, Junjun Jiang, Xianming Liu, Guanglu Sun, Yi Xiao
Abstract
Rain significantly degrades the performance of computer vision systems, particularly in applications like autonomous driving and video surveillance. While existing deraining methods have made considerable progress, they often struggle with fidelity of semantic and spatial details. To address these limitations, we propose the Multi-Prior Hierarchical Mamba (MPHM) network for image deraining. This novel architecture synergistically integrates macro-semantic textual priors (CLIP) for task-level semantic guidance and micro-structural visual priors (DINOv2) for scene-aware structural information. To alleviate potential conflicts between heterogeneous priors, we devise a progressive Priors Fusion Injection (PFI) that strategically injects complementary cues at different decoder levels. Meanwhile, we equip the backbone network with an elaborate Hierarchical Mamba Module (HMM) to facilitate robust feature representation, featuring a Fourier-enhanced dual-path design that concurrently addresses global context modeling and local detail recovery. Comprehensive experiments demonstrate MPHM's state-of-the-art performance, achieving a 0.57 dB PSNR gain on the Rain200H dataset while delivering superior generalization on real-world rainy scenarios.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b11f565c-3cd9-468a-8d7b-07fc80470bf7Builds on18
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Restormer: Efficient Transformer for High-Resolution Image RestorationSyed Waqas Zamir, Aditya Arora, Salman Khan, Munawar Hayat et al.CVPR 2022 · 3,348 citations
- VMamba: Visual State Space ModelYue Liu, Yunjie Tian, Yuzhong Zhao, Hongtian Yu et al.NeurIPS 2024 · 3,199 citations
- Uformer: A General U-Shaped Transformer for Image RestorationZhendong Wang, Xiaodong Cun, Jianmin Bao, Wengang Zhou et al.CVPR 2022 · 1,970 citations
- Structure-Preserving Deraining with Residue Channel Prior GuidanceQiaosi Yi, Juncheng Li, Qinyan Dai, Faming Fang et al.ICCV 2021 · 159 citations
Related papers
- FreqMamba: Viewing Mamba from a Frequency Perspective for Image DerainingZhen Zou, Hu Yu, Jie Huang, Feng ZhaoACM MM 2024 · 73 citations
- PRE-Mamba: A 4D State Space Model for Ultra-High-Frequent Event Camera DerainingCiyu Ruan, Ruishan Guo, Zihang Gong, Jingao Xu et al.ICCV 2025 · 3 citations
- Close the Loop: A Unified Bottom-Up and Top-Down Paradigm for Joint Image Deraining and SegmentationYi Li, Yi Chang, Changfeng Yu, Luxin YanAAAI 2022 · 31 citations
- Multi-Scale Progressive Fusion Network for Single Image DerainingKui Jiang, Zhongyuan Wang, Peng Yi, Chen Chen et al.CVPR 2020
- Multi-Decoding Deraining Network and Quasi-Sparsity Based TrainingYinglong Wang, Chao Ma, Bing ZengCVPR 2021
