Exploiting Diffusion Prior for Generalizable Dense Prediction
Hsin-Ying Lee, Hung-Yu Tseng, Hsin-Ying Lee, Ming-Hsuan Yang
摘要
Contents generated by recent advanced Text-to-Image (T2I) diffusion models are sometimes too imaginative for existing off-the-shelf dense predictors to estimate due to the immitigable domain gap. We introduce DMP, a pipeline utilizing pre-trained T2I models as a prior for dense prediction tasks. To address the misalignment between deterministic prediction tasks and stochastic T2I models, we reformulate the diffusion process through a sequence of interpo-lations, establishing a deterministic mapping between input RGB images and output prediction distributions. To preserve generalizability, we use low-rank adaptation to fine-tune pre-trained models. Extensive experiments across five tasks, including 3D property estimation, semantic segmentation, and intrinsic image decomposition, showcase the efficacy of the proposed method. Despite limited-domain training data, the approach yields faithful estimations for arbitrary images, surpassing existing state-of-the-art algorithms. The code is available at https://github.com/shinying/dmp.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper19
- DICEPTION: A Generalist Diffusion Model for Visual Perceptual TasksCanyu Zhao, Yanlong Sun, Mingyu Liu, Huanyi Zheng 等NeurIPS 2025 · 被引用 45 次
- Unleashing the Potential of the Diffusion Model in Few-shot Semantic SegmentationMuzhi Zhu, Yang Liu, Zekai Luo, Chenchen Jing 等NeurIPS 2024 · 被引用 31 次
- Zero-shot Depth Completion via Test-time Alignment with Affine-invariant Depth PriorLee Hyoseok, Kyeong Seon Kim, Byung-Ki Kwon, Tae-Hyun OhAAAI 2025 · 被引用 11 次
- GlassWizard: Harvesting Diffusion Priors for Glass Surface DetectionWenxue Li, Tian Ye, Xinyu Xiong, Jinbin Bai 等ICCV 2025 · 被引用 8 次
- LightCity: An Urban Dataset for Outdoor Inverse Rendering and Reconstruction Under Multi-Illumination ConditionsJingjing Wang, Qirui Hu, Chong Bao, Yuke Zhu 等ICCV 2025 · 被引用 5 次
它引用的顶会 Paper41
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 被引用 6,759 次
相关 Paper
- DDP: Diffusion Model for Dense Visual PredictionYuanfeng Ji, Zhe Chen, Enze Xie, Lanqing Hong 等ICCV 2023 · 被引用 223 次
- DNF-Intrinsic: Deterministic Noise-Free Diffusion for Indoor Inverse RenderingRongjia Zheng, Qing Zhang, Chengjiang Long, Wei-Shi ZhengICCV 2025 · 被引用 2 次
- Sketch-Guided Text-to-Image Diffusion ModelsAndrey Voynov, Kfir Aberman, Daniel Cohen-OrSIGGRAPH 2023 · 被引用 168 次
- DRIP: Unleashing Diffusion Priors for Joint Foreground and Alpha Prediction in Image MattingXiaodi Li, Zongxin Yang, Ruijie Quan, Yi YangNeurIPS 2024 · 被引用 15 次
- DPMesh: Exploiting Diffusion Prior for Occluded Human Mesh RecoveryYixuan Zhu, Ao Li, Yansong Tang, Wenliang Zhao 等CVPR 2024 · 被引用 10 次
