Score Distillation via Reparametrized DDIM
Artem Lukoianov, Haitz Sáez de Ocáriz Borde, Kristjan H. Greenewald, Vitor Guizilini, Timur M. Bagautdinov, Vincent Sitzmann, Justin M. Solomon
摘要
While 2D diffusion models generate realistic, high-detail images, 3D shape generation methods like Score Distillation Sampling (SDS) built on these 2D diffusion models produce cartoon-like, over-smoothed shapes. To help explain this discrepancy, we show that the image guidance used in Score Distillation can be understood as the velocity field of a 2D denoising generative process, up to the choice of a noise term. In particular, after a change of variables, SDS resembles a high-variance version of Denoising Diffusion Implicit Models (DDIM) with a differently-sampled noise term: SDS introduces noise i.i.d. randomly at each step, while DDIM infers it from the previous noise predictions. This excessive variance can lead to over-smoothing and unrealistic outputs. We show that a better noise approximation can be recovered by inverting DDIM in each SDS update step. This modification makes SDS's generative process for 2D images almost identical to DDIM. In 3D, it removes over-smoothing, preserves higher-frequency detail, and brings the generation quality closer to that of 2D samplers. Experimentally, our method achieves better or similar 3D generation quality compared to other state-of-the-art Score Distillation methods, all without training additional neural networks or multi-view supervision, and providing useful insights into relationship between 2D and 3D asset generation with diffusion models.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper23
- Rethinking Score Distillation as a Bridge Between Image DistributionsDavid McAllister, Songwei Ge, Jia-Bin Huang, David Jacobs 等NeurIPS 2024 · 被引用 43 次
- Rectified CFG++ for Flow Based ModelsShreshth Saini, Shashank Gupta, Alan BovikNeurIPS 2025 · 被引用 20 次
- SCoT: Unifying Consistency Models and Rectified Flows via Straight-Consistent TrajectoriesZhangkai Wu, Xuhui Fan, Hongyu Wu, Longbing CaoNeurIPS 2025 · 被引用 12 次
- Generating Physically Stable and Buildable Brick Structures from TextAva Pun, Kangle Deng, Ruixuan Liu, Deva Ramanan 等ICCV 2025 · 被引用 9 次
- DiffuMatch: Category-Agnostic Spectral Diffusion Priors for Robust Non-Rigid Shape MatchingEmery Pierson, Lei Li, Angela Dai, Maks OvsjanikovICCV 2025 · 被引用 5 次
它引用的顶会 Paper31
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
相关 Paper
- Consistent Flow Distillation for Text-to-3D GenerationRunjie Yan, Yinbo Chen, Xiaolong WangICLR 2025
- Rethinking Score Distilling Sampling for 3D Editing and GenerationXingyu Miao, Haoran Duan, Yang Long, Jungong HanICML 2025
- Diffusion Time-step Curriculum for One Image to 3D GenerationXuanyu Yi, Zike Wu, Qingshan Xu, Pan Zhou 等CVPR 2024 · 被引用 7 次
- Noise-free Score DistillationOren Katzir, Or Patashnik, Daniel Cohen-Or, Dani LischinskiICLR 2024 · 被引用 101 次
- Consistent3D: Towards Consistent High-Fidelity Text-to-3D Generation with Deterministic Sampling PriorZike Wu, Pan Zhou, Xuanyu Yi, Xiaoding Yuan 等CVPR 2024 · 被引用 15 次
