Paint-it: Text-to-Texture Synthesis via Deep Convolutional Texture Map Optimization and Physically-Based Rendering
Kim Youwang, Tae-Hyun Oh, Gerard Pons-Moll
Abstract
We present Paint-it, a text-driven high-fidelity texture map synthesis method for 3D meshes via neural re-parameterized texture optimization. Paint-it synthesizes texture maps from a text description by synthesis-through-optimization, exploiting the Score-Distillation Sampling (SDS). We observe that directly applying SDS yields undesirable texture quality due to its noisy gradients. We reveal the importance of texture parameterization when using SDS. Specifically, we propose Deep Convolutional Physically-Based Rendering (DC-PBR) parameterization, which re-parameterizes the physically-based rendering (PBR) texture maps with randomly initialized convolution-based neural kernels, instead of a standard pixel-based parameterization. We show that DC-PBR inherently schedules the optimization curriculum according to texture frequency and naturally filters out the noisy signals from SDS. In experiments, Paint-it obtains remarkable quality PBR texture maps within 15 min., given only a text description. We demonstrate the generalizability and practicality of Paint-it by synthesizing high-quality texture maps for large-scale mesh datasets and showing test-time applications such as relighting and material control using a popular graphics engine. Project page: https://kim-youwang.github.io/paint-it.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers37
- Human-3Diffusion: Realistic Avatar Creation via Explicit 3D Consistent Diffusion ModelsYuxuan Xue, Xianghui Xie, Riccardo Marin, Gerard Pons-MollNeurIPS 2024 · 49 citations
- Rethinking Score Distillation as a Bridge Between Image DistributionsDavid McAllister, Songwei Ge, Jia-Bin Huang, David Jacobs et al.NeurIPS 2024 · 43 citations
- DreamMat: High-quality PBR Material Generation with Geometry- and Light-aware Diffusion ModelsYuqing Zhang, Yuan Liu, Zhiyu Xie, Lei Yang et al.SIGGRAPH 2024 · 28 citations
- SyncTweedies: A General Generative Framework Based on Synchronized DiffusionsJaihoon Kim, Juil Koo, Kyeongmin Yeo, Minhyuk SungNeurIPS 2024 · 27 citations
- Make-it-Real: Unleashing Large Multimodal Model for Painting 3D Objects with Realistic MaterialsYe Fang, Zeyi Sun, Tong Wu, Jiaqi Wang et al.NeurIPS 2024 · 21 citations
Builds on34
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li et al.NeurIPS 2022 · 8,965 citations
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 4,089 citations
- Consistency ModelsYang Song, Prafulla Dhariwal, Mark Chen, Ilya SutskeverICML 2023 · 1,720 citations
Related papers
- SceneTex: High-Quality Texture Synthesis for Indoor Scenes via Diffusion PriorsDave Zhenyu Chen, Haoxuan Li, Hsin-Ying Lee, Sergey Tulyakov et al.CVPR 2024
- TextureDreamer: Image-Guided Texture Synthesis through Geometry-Aware DiffusionYu-Ying Yeh, Jia-Bin Huang, Changil Kim, Lei Xiao et al.CVPR 2024 · 31 citations
- ProlificDreamer: High-Fidelity and Diverse Text-to-3D Generation with Variational Score DistillationZhengyi Wang, Cheng Lu, Yikai Wang, Fan Bao et al.NeurIPS 2023 · 1,498 citations
- IntrinsiX: High-Quality PBR Generation using Image PriorsPeter Kocsis, Lukas Höllein, Matthias NießnerNeurIPS 2025 · 20 citations
- Text2VDM: Text to Vector Displacement Maps for Expressive and Interactive 3D SculptingHengyu Meng, Duotun Wang, Zhijing Shao, Ligang Liu et al.ICCV 2025 · 2 citations
