PlacidDreamer: Advancing Harmony in Text-to-3D Generation
Shuo Huang, Shikun Sun, Zixuan Wang, Xiaoyu Qin, Yanmin Xiong, Yuan Zhang, Pengfei Wan, Di Zhang, Jia Jia
摘要
Recently, text-to-3D generation has attracted significant attention, resulting in notable performance enhancements. Previous methods utilize end-to-end 3D generation models to initialize 3D Gaussians, multi-view diffusion models to enforce multi-view consistency, and text-to-image diffusion models to refine details with score distillation algorithms. However, these methods exhibit two limitations. Firstly, they encounter conflicts in generation directions since different models aim to produce diverse 3D assets. Secondly, the issue of over-saturation in score distillation has not been thoroughly investigated and solved. To address these limitations, we propose PlacidDreamer, a text-to-3D framework that harmonizes initialization, multi-view generation, and text-conditioned generation with a single multi-view diffusion model, while simultaneously employing a novel score distillation algorithm to achieve balanced saturation. To unify the generation direction, we introduce the Latent-Plane module, a training-friendly plug-in extension that enables multi-view diffusion models to provide fast geometry reconstruction for initialization and enhanced multi-view images to personalize the text-to-image diffusion model. To address the over-saturation problem, we propose to view score distillation as a multi-objective optimization problem and introduce the Balanced Score Distillation algorithm, which offers a Pareto Optimal solution that achieves both rich details and balanced saturation. Extensive experiments validate the outstanding capabilities of our PlacidDreamer. The code is available at https://github.com/HansenHuang0823/PlacidDreamer.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Identity Preserving 3D Head Stylization with Multiview Score DistillationBahri Batuhan Bilecen, Ahmet Berke Gokmen, Furkan Guzelant, Aysegul DundarICCV 2025 · 被引用 3 次
- Inferring Compositional 4D Scenes without Ever Seeing OneAhmet Berke Gökmen, Ajad Chhatkuli, Luc Van Gool, Danda PaudelCVPR 2026 · 被引用 1 次
- CoGrad3D: Spatially-Coupled Timestep Optimization with Orthogonal Gradient Fusion for 3D GenerationHaoyang Tong, Hongbo Wang, Jin Liu, Qi Wang 等AAAI 2026
- Progressive Rendering Distillation: Adapting Stable Diffusion for Instant Text-to-Mesh Generation without 3D DataZhiyuan Ma, Xinyue Liang, Rongyuan Wu, Xiangyu Zhu 等CVPR 2025
- CAdam: Context-Adaptive Moment Estimation for 3D Gaussian Densification in Generative DistillationSeungJeh Chung, Geonho Park, Misong Kim, HyeongYeop KangSIGGRAPH 2026
它引用的顶会 Paper45
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
相关 Paper
- ProlificDreamer: High-Fidelity and Diverse Text-to-3D Generation with Variational Score DistillationZhengyi Wang, Cheng Lu, Yikai Wang, Fan Bao 等NeurIPS 2023 · 被引用 1,498 次
- Retrieval-Augmented Score Distillation for Text-to-3D GenerationJunyoung Seo, Susung Hong, Wooseok Jang, Inès Hyeonsu Kim 等ICML 2024 · 被引用 14 次
- Debiasing Scores and Prompts of 2D Diffusion for View-consistent Text-to-3D GenerationSusung Hong, Donghoon Ahn, Seungryong KimNeurIPS 2023 · 被引用 46 次
- HandDreamer: Zero-Shot Text to 3D Hand Model Generation using Corrective Hand Shape GuidanceGreen Rosh, Prateek Kukreja, Vishakha SR, Pawan Prasad B HCVPR 2026
- A3D: Does Diffusion Dream about 3D Alignment?Savva Victorovich Ignatyev, Nina Konovalova, Daniil Selikhanovych, Oleg Voynov 等ICLR 2025
