Score Distillation via Reparametrized DDIM
Artem Lukoianov, Haitz Sáez de Ocáriz Borde, Kristjan H. Greenewald, Vitor Guizilini, Timur M. Bagautdinov, Vincent Sitzmann, Justin M. Solomon
Abstract
While 2D diffusion models generate realistic, high-detail images, 3D shape generation methods like Score Distillation Sampling (SDS) built on these 2D diffusion models produce cartoon-like, over-smoothed shapes. To help explain this discrepancy, we show that the image guidance used in Score Distillation can be understood as the velocity field of a 2D denoising generative process, up to the choice of a noise term. In particular, after a change of variables, SDS resembles a high-variance version of Denoising Diffusion Implicit Models (DDIM) with a differently-sampled noise term: SDS introduces noise i.i.d. randomly at each step, while DDIM infers it from the previous noise predictions. This excessive variance can lead to over-smoothing and unrealistic outputs. We show that a better noise approximation can be recovered by inverting DDIM in each SDS update step. This modification makes SDS's generative process for 2D images almost identical to DDIM. In 3D, it removes over-smoothing, preserves higher-frequency detail, and brings the generation quality closer to that of 2D samplers. Experimentally, our method achieves better or similar 3D generation quality compared to other state-of-the-art Score Distillation methods, all without training additional neural networks or multi-view supervision, and providing useful insights into relationship between 2D and 3D asset generation with diffusion models.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3bf5bcae-61e7-4bd9-a99b-30e6f23cbcedCited by top-tier papers23
- Rethinking Score Distillation as a Bridge Between Image DistributionsDavid McAllister, Songwei Ge, Jia-Bin Huang, David Jacobs et al.NeurIPS 2024 · 43 citations
- Rectified CFG++ for Flow Based ModelsShreshth Saini, Shashank Gupta, Alan BovikNeurIPS 2025 · 20 citations
- SCoT: Unifying Consistency Models and Rectified Flows via Straight-Consistent TrajectoriesZhangkai Wu, Xuhui Fan, Hongyu Wu, Longbing CaoNeurIPS 2025 · 12 citations
- Generating Physically Stable and Buildable Brick Structures from TextAva Pun, Kangle Deng, Ruixuan Liu, Deva Ramanan et al.ICCV 2025 · 9 citations
- DiffuMatch: Category-Agnostic Spectral Diffusion Priors for Robust Non-Rigid Shape MatchingEmery Pierson, Lei Li, Angela Dai, Maks OvsjanikovICCV 2025 · 5 citations
Builds on31
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
Related papers
- Consistent Flow Distillation for Text-to-3D GenerationRunjie Yan, Yinbo Chen, Xiaolong WangICLR 2025
- Rethinking Score Distilling Sampling for 3D Editing and GenerationXingyu Miao, Haoran Duan, Yang Long, Jungong HanICML 2025
- Diffusion Time-step Curriculum for One Image to 3D GenerationXuanyu Yi, Zike Wu, Qingshan Xu, Pan Zhou et al.CVPR 2024 · 7 citations
- Noise-free Score DistillationOren Katzir, Or Patashnik, Daniel Cohen-Or, Dani LischinskiICLR 2024 · 101 citations
- Consistent3D: Towards Consistent High-Fidelity Text-to-3D Generation with Deterministic Sampling PriorZike Wu, Pan Zhou, Xuanyu Yi, Xiaoding Yuan et al.CVPR 2024 · 15 citations
