Sin3DM: Learning a Diffusion Model from a Single 3D Textured Shape
Rundi Wu, Ruoshi Liu, Carl Vondrick, Changxi Zheng
Abstract
Synthesizing novel 3D models that resemble the input example has long been pursued by graphics artists and machine learning researchers. In this paper, we present Sin3DM, a diffusion model that learns the internal patch distribution from a single 3D textured shape and generates high-quality variations with fine geometry and texture details. Training a diffusion model directly in 3D would induce large memory and computational cost. Therefore, we first compress the input into a lower-dimensional latent space and then train a diffusion model on it. Specifically, we encode the input 3D textured shape into triplane feature maps that represent the signed distance and texture fields of the input. The denoising network of our diffusion model has a limited receptive field to avoid overfitting, and uses triplane-aware 2D convolution blocks to improve the result quality. Aside from randomly generating new samples, our model also facilitates applications such as retargeting, outpainting and local editing. Through extensive qualitative and quantitative evaluation, we show that our method outperforms prior methods in generation quality of 3D shapes.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 03d2391f-850b-4cf2-b941-4d4aa1bb5dc3Cited by top-tier papers13
- pix2gestalt: Amodal Segmentation by Synthesizing WholesEge Ozguroglu, Ruoshi Liu, Dídac Surís, Dian Chen et al.CVPR 2024 · 24 citations
- ThemeStation: Generating Theme-Aware 3D Assets from Few ExemplarsZhenwei Wang, Tengfei Wang, Gerhard P. Hancke, Ziwei Liu et al.SIGGRAPH 2024 · 7 citations
- CodecNeRF: Toward Fast Encoding and Decoding, Compact, and High-quality Novel-view SynthesisGyeongjin Kang, Younggeun Lee, Seungjun Oh, Eunbyung ParkAAAI 2025 · 5 citations
- UV-free Texture Generation with Denoising and Geodesic Heat DiffusionSimone Foti, Stefanos Zafeiriou, Tolga BirdalNeurIPS 2024 · 5 citations
- Single Mesh Diffusion Models with Field Latents for Texture GenerationThomas W. Mitchel, Carlos Esteves, Ameesh MakadiaCVPR 2024 · 4 citations
Builds on36
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li et al.NeurIPS 2022 · 8,965 citations
Related papers
- TEXTure: Text-Guided Texturing of 3D ShapesElad Richardson, Gal Metzer, Yuval Alaluf, Raja Giryes et al.SIGGRAPH 2023 · 196 citations
- TriTex: Learning Texture from a Single Mesh via Triplane Semantic FeaturesDana Cohen-Bar, Daniel Cohen-Or, Gal Chechik, Yoni KastenCVPR 2025
- Single Motion DiffusionSigal Raab, Inbal Leibovitch, Guy Tevet, Moab Arar et al.ICLR 2024 · 81 citations
- Diffusion Texture PaintingAnita Hu, Nishkrit Desai, Hassan Abu Alhaija, Seung Wook Kim et al.SIGGRAPH 2024 · 16 citations
- UDiFF: Generating Conditional Unsigned Distance Fields with Optimal Wavelet DiffusionJunsheng Zhou, Weiqi Zhang, Baorui Ma, Kanle Shi et al.CVPR 2024 · 11 citations
