Exploiting Diffusion Prior for Generalizable Dense Prediction
Hsin-Ying Lee, Hung-Yu Tseng, Hsin-Ying Lee, Ming-Hsuan Yang
Abstract
Contents generated by recent advanced Text-to-Image (T2I) diffusion models are sometimes too imaginative for existing off-the-shelf dense predictors to estimate due to the immitigable domain gap. We introduce DMP, a pipeline utilizing pre-trained T2I models as a prior for dense prediction tasks. To address the misalignment between deterministic prediction tasks and stochastic T2I models, we reformulate the diffusion process through a sequence of interpo-lations, establishing a deterministic mapping between input RGB images and output prediction distributions. To preserve generalizability, we use low-rank adaptation to fine-tune pre-trained models. Extensive experiments across five tasks, including 3D property estimation, semantic segmentation, and intrinsic image decomposition, showcase the efficacy of the proposed method. Despite limited-domain training data, the approach yields faithful estimations for arbitrary images, surpassing existing state-of-the-art algorithms. The code is available at https://github.com/shinying/dmp.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers19
- DICEPTION: A Generalist Diffusion Model for Visual Perceptual TasksCanyu Zhao, Yanlong Sun, Mingyu Liu, Huanyi Zheng et al.NeurIPS 2025 · 45 citations
- Unleashing the Potential of the Diffusion Model in Few-shot Semantic SegmentationMuzhi Zhu, Yang Liu, Zekai Luo, Chenchen Jing et al.NeurIPS 2024 · 31 citations
- Zero-shot Depth Completion via Test-time Alignment with Affine-invariant Depth PriorLee Hyoseok, Kyeong Seon Kim, Byung-Ki Kwon, Tae-Hyun OhAAAI 2025 · 11 citations
- GlassWizard: Harvesting Diffusion Priors for Glass Surface DetectionWenxue Li, Tian Ye, Xinyu Xiong, Jinbin Bai et al.ICCV 2025 · 8 citations
- LightCity: An Urban Dataset for Outdoor Inverse Rendering and Reconstruction Under Multi-Illumination ConditionsJingjing Wang, Qirui Hu, Chong Bao, Yuke Zhu et al.ICCV 2025 · 5 citations
Builds on41
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li et al.NeurIPS 2022 · 8,965 citations
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 6,759 citations
Related papers
- DDP: Diffusion Model for Dense Visual PredictionYuanfeng Ji, Zhe Chen, Enze Xie, Lanqing Hong et al.ICCV 2023 · 223 citations
- DNF-Intrinsic: Deterministic Noise-Free Diffusion for Indoor Inverse RenderingRongjia Zheng, Qing Zhang, Chengjiang Long, Wei-Shi ZhengICCV 2025 · 2 citations
- Sketch-Guided Text-to-Image Diffusion ModelsAndrey Voynov, Kfir Aberman, Daniel Cohen-OrSIGGRAPH 2023 · 168 citations
- DRIP: Unleashing Diffusion Priors for Joint Foreground and Alpha Prediction in Image MattingXiaodi Li, Zongxin Yang, Ruijie Quan, Yi YangNeurIPS 2024 · 15 citations
- DPMesh: Exploiting Diffusion Prior for Occluded Human Mesh RecoveryYixuan Zhu, Ao Li, Yansong Tang, Wenliang Zhao et al.CVPR 2024 · 10 citations
