Rethinking Score Distillation as a Bridge Between Image Distributions
David McAllister, Songwei Ge, Jia-Bin Huang, David Jacobs, Alexei A. Efros, Aleksander Holynski, Angjoo Kanazawa
摘要
Score distillation sampling (SDS) has proven to be an important tool, enabling the use of large-scale diffusion priors for tasks operating in data-poor domains. Unfortunately, SDS has a number of characteristic artifacts that limit its usefulness in general-purpose applications. In this paper, we make progress toward understanding the behavior of SDS and its variants by viewing them as solving an optimal-cost transport path from a source distribution to a target distribution. Under this new interpretation, these methods seek to transport corrupted images (source) to the natural image distribution (target). We argue that current methods' characteristic artifacts are caused by (1) linear approximation of the optimal path and (2) poor estimates of the source distribution. We show that calibrating the text conditioning of the source distribution can produce high-quality generation and translation results with little extra overhead. Our method can be easily applied across many domains, matching or beating the performance of specialized methods. We demonstrate its utility in text-to-2D, text-based NeRF optimization, translating paintings to real images, optical illusion generation, and 3D sketch-to-real. We compare our method to existing approaches for score distillation sampling and show that it can produce high-frequency details with realistic colors.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper22
- Generating Physically Stable and Buildable Brick Structures from TextAva Pun, Kangle Deng, Ruixuan Liu, Deva Ramanan 等ICCV 2025 · 被引用 9 次
- Coupled Diffusion Sampling for Training-Free Multi-View Image EditingHadi Alzayer, Yunzhi Zhang, Chen Geng, Jia-Bin Huang 等CVPR 2026 · 被引用 6 次
- Let it Snow! Animating 3D Gaussian Scenes with Dynamic Weather Effects via Physics-Guided Score DistillationGal Fiebelman, Hadar Averbuch-Elor, Sagie BenaimCVPR 2026 · 被引用 6 次
- Delta Rectified Flow Sampling for Text-to-Image EditingGaspard Beaudouin, Minghan Li, Jaeyeon Kim, Sung-Hoon Yoon 等CVPR 2026 · 被引用 4 次
- Efficient Autoregressive Shape Generation Via Octree-Based Adaptive TokenizationKangle Deng, Hsueh-Ti Derek Liu, Yiheng Zhu, Xiaoxia Sun 等ICCV 2025 · 被引用 4 次
它引用的顶会 Paper51
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
相关 Paper
- Delta Denoising ScoreAmir Hertz, Kfir Aberman, Daniel Cohen-OrICCV 2023 · 被引用 136 次
- Noise-free Score DistillationOren Katzir, Or Patashnik, Daniel Cohen-Or, Dani LischinskiICLR 2024 · 被引用 101 次
- Target-Balanced Score DistillationZhou Xu, Qi Wang, Yuxiao Yang, Luyuan Zhang 等AAAI 2026
- AnchorDS: Anchoring Dynamic Sources for Semantically Consistent Text-to-3D GenerationJiayin Zhu, Linlin Yang, Yicong Li, Angela YaoAAAI 2026 · 被引用 2 次
- Consistent3D: Towards Consistent High-Fidelity Text-to-3D Generation with Deterministic Sampling PriorZike Wu, Pan Zhou, Xuanyu Yi, Xiaoding Yuan 等CVPR 2024 · 被引用 15 次
