Identity-preserving Distillation Sampling by Fixed-Point Iterator
Seonhwa Kim, Jiwon Kim, Soobin Park, Donghoon Ahn, Jiwon Kang, Seungryong Kim, Kyong Hwan Jin, Eunju Cha
摘要
Score distillation sampling (SDS) demonstrates a powerful capability for text-conditioned 2D image and 3D object generation by distilling the knowledge from learned score functions. However, SDS often suffers from blurriness caused by noisy gradients. When SDS meets the image editing, such degradations can be reduced by adjusting bias shifts using reference pairs, but the de-biasing techniques are still corrupted by erroneous gradients. To this end, we introduce Identity-preserving Distillation Sampling (IDS), which compensates for the gradient leading to undesired changes in the results. Based on the analysis that these errors come from the text-conditioned scores, a new regularization technique, called fixed-point iterative regularization (FPR), is proposed to modify the score itself, driving the preservation of the identity even including poses and structures. Thanks to a self-correction by FPR, the proposed method provides clear and unambiguous representations corresponding to the given prompts in image-to-image editing and editable neural radiance field (NeRF). The structural consistency between the source and the edited data is obviously maintained compared to other state-ofthe-art methods. Our code is https://github.com/ shhh0620/IDS
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Dual Recursive Feedback on Generation and Appearance Latents for Pose-Robust Text-to-Image DiffusionJiwon Kim, Pu-Reum Kim, Seonhwa Kim, Soobin Park 等ICCV 2025
- Semantic Alignment for Pose-Invariant Identity Preserving DiffusionJiwon Kim, Seonhwa Kim, Soobin Park, Eunju Cha 等CVPR 2026
它引用的顶会 Paper20
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
相关 Paper
- ED-NeRF: Efficient Text-Guided Editing of 3D Scene With Latent Space NeRFJangho Park, Gihyun Kwon, Jong Chul YeICLR 2024 · 被引用 22 次
- Decorate3D: Text-Driven High-Quality Texture Generation for Mesh Decoration in the WildYanhui Guo, Xinxin Zuo, Peng Dai, Juwei Lu 等NeurIPS 2023 · 被引用 13 次
- Stable Score DistillationHaiming Zhu, Yangyang Xu, Chenshu Xu, Tingrui Shen 等ICCV 2025 · 被引用 2 次
- Rethinking Score Distilling Sampling for 3D Editing and GenerationXingyu Miao, Haoran Duan, Yang Long, Jungong HanICML 2025
- Perturb-and-Revise: Flexible 3D Editing with Generative TrajectoriesSusung Hong, Johanna Suvi Karras, Ricardo Martin-Brualla, Ira Kemelmacher-ShlizermanCVPR 2025
