Prompt-SID: Learning Structural Representation Prompt via Latent Diffusion for Single Image Denoising
Huaqiu Li, Wang Zhang, Xiaowan Hu, Tao Jiang, Zikang Chen, Haoqian Wang
Abstract
Many studies have concentrated on constructing supervised models utilizing paired datasets for image denoising, which proves to be expensive and time-consuming. Current self-supervised and unsupervised approaches typically rely on blind-spot networks or sub-image pairs sampling, resulting in pixel information loss and destruction of detailed structural information, thereby significantly constraining the efficacy of such methods. In this paper, we introduce Prompt-SID, a prompt-learning-based single image denoising framework that emphasizes the preservation of structural details. This approach is trained in a self-supervised manner using downsampled image pairs. It captures original-scale image information through structural encoding and integrates this prompt into the denoiser. To achieve this, we propose a structural representation generation model based on the latent diffusion process and design a structural attention module within the transformer-based denoiser architecture to decode the prompt. Additionally, we introduce a scale replay training mechanism, which effectively mitigates the scale gap from images of different resolutions. We conduct comprehensive experiments on synthetic, real-world, and fluorescence imaging datasets, showcasing the remarkable effectiveness of Prompt-SID.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers4
- Does Your Reasoning Model Implicitly Know When to Stop Thinking?Zixuan Huang, Xin Xia, Yuxi Ren, Jianbin Zheng et al.ICML 2026 · 21 citations
- Real-Time Aligned Reward Model beyond SemanticsZixuan Huang, Xin Xia, Yuxi Ren, Jianbin Zheng et al.ICML 2026 · 18 citations
- LD-RPS: Zero-Shot Unified Image Restoration via Latent Diffusion Recurrent Posterior SamplingHuaqiu Li, Yong Wang, Tongwen Huang, Hailang Huang et al.ICCV 2025 · 4 citations
- Zero-Shot Blind-Spot Image Denoising via Cross-Scale Non-Local Pixel RefillingQilong Guo, Tianjing Zhang, Zhiyuan Ma, Hui JiNeurIPS 2025
Builds on14
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
- Restormer: Efficient Transformer for High-Resolution Image RestorationSyed Waqas Zamir, Aditya Arora, Salman Khan, Munawar Hayat et al.CVPR 2022 · 3,348 citations
- Denoising Diffusion Restoration ModelsBahjat Kawar, Michael Elad, Stefano Ermon, Jiaming SongNeurIPS 2022 · 1,439 citations
- Real Image Denoising With Feature AttentionSaeed Anwar, Nick BarnesICCV 2019 · 644 citations
Related papers
- Next-Scale Prediction: A Self-Supervised Approach for Real-World Image DenoisingYiwen Shan, Haiyu Zhao, Peng Hu, Xi Peng et al.CVPR 2026 · 2 citations
- Iterative Denoiser and Noise Estimator for Self-Supervised Image DenoisingYunhao Zou, Chenggang Yan, Ying FuICCV 2023 · 30 citations
- SinDDM: A Single Image Denoising Diffusion ModelVladimir Kulikov, Shahar Yadin, Matan Kleiner, Tomer MichaeliICML 2023 · 113 citations
- Pseudo-Siamese Blind-spot Transformers for Self-Supervised Real-World DenoisingYuhui Quan, Tianxiang Zheng, Hui JiNeurIPS 2024 · 5 citations
- Zero-Shot Noise2Mean: Gap Minimization for Efficient Denoising from a Single Noisy ImageDuo Liu, Yiqi Shi, Guoyin Zhang, Sizhao Li et al.AAAI 2025
