From One to More: Contextual Part Latents for 3D Generation
Shaocong Dong, Lihe Ding, Xiao Chen, Yaokun Li, Yuxin Wang, Yucheng Wang, Qi Wang, Jaehyeok Kim, Chenjian Gao, Zhanpeng Huang, Zibin Wang, Tianfan Xue, Dan Xu
Abstract
To generate 3D objects, early research focused on multi-view-driven approaches relying solely on 2D renderings. Recently, the 3D native latent diffusion paradigm has demonstrated superior performance in 3D generation, because it fully leverages the geometric information provided in ground truth 3D data. Despite its fast development, 3D diffusion still faces three challenges. First, the majority of these methods represent a 3D object by one single latent, regardless of its complexity. This may lead to detail loss when generating 3D objects with multiple complicated parts. Second, most 3D assets are designed parts by parts, yet the current holistic latent representation overlooks the independence of these parts and their interrelationships, limiting the model's generative ability. Third, current methods rely on global conditions (e.g., text, image, point cloud) to control the generation process, lacking detailed controllability. Therefore, motivated by how 3D designers create a 3D object, we present a new part-based 3D generation framework, CoPart, which represents a 3D object with multiple contextual part latents and simultaneously generates coherent 3D parts. This part-based framework has several advantages, including: i) reduces the encoding burden of intricate objects by decomposing them into simpler parts, ii) facilitates part learning and part relationship modeling, and iii) naturally supports part-level control. Furthermore, to ensure the coherence of part latents and to harness the powerful priors from foundation models, we propose a novel mutual guidance strategy to fine-tune pre-trained diffusion models for joint part latent denoising. Benefiting from the part-based representation, we demonstrate that CoPart can support various applications including part-editing, articulated object generation, and mini-scene generation. Moreover, we collect a new large-scale 3D part dataset named Partverse from Objaverse through automatic mesh segmentation and subsequent human post-annotations. By training on the proposed dataset, CoPart achieves promising part-based 3D generation with high controllability. Project page: https://hkdsc.github.io/project/copart.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 887b3f0a-1f92-493e-b0f7-1964ec3df585Cited by top-tier papers11
- FullPart: Generating each 3D Part at Full ResolutionLihe Ding, Shaocong Dong, Yaokun Li, Chenjian Gao et al.ICLR 2026 · 17 citations
- PhysForge: Generating Physics-Grounded 3D Assets for Interactive Virtual WorldYunhan Yang, Chunshi Wang, Junliang Ye, YANG LI et al.ICML 2026 · 6 citations
- M3DLayout: A Multi-Source Dataset of 3D Indoor Layouts and Structured Descriptions for 3D GenerationYiheng Zhang, Zhuojiang Cai, Mingdao Wang, Meitong Guo et al.CVPR 2026 · 5 citations
- Repurposing 3D Generative Model for Autoregressive Layout GenerationHaoran Feng, Yifan Niu, Zehuan Huang, Yangtian Sun et al.CVPR 2026 · 3 citations
- JRM: Joint Reconstruction Model for Multiple Objects without AlignmentQirui Wu, Mohd Yawar Nihal Siddiqui, Duncan Frost, Samir Aroudj et al.CVPR 2026 · 2 citations
Builds on34
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao et al.ICCV 2023 · 13,211 citations
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 6,759 citations
- Scalable Diffusion Models with TransformersWilliam Peebles, Saining XieICCV 2023 · 5,568 citations
- Zero-1-to-3: Zero-shot One Image to 3D ObjectRuoshi Liu, Rundi Wu, Basile Van Hoorick, Pavel Tokmakov et al.ICCV 2023 · 1,662 citations
Related papers
- UniPart: Part-Level 3D Generation with Unified 3D Geom-Seg LatentsXufan He, Yushuang Wu, Xiaoyang Guo, Chongjie Ye et al.CVPR 2026 · 9 citations
- HoloPart: Generative 3D Part Amodal SegmentationYunhan Yang, Yuanchen Guo, Yukun Huang, Zi-Xin Zou et al.ICLR 2026 · 62 citations
- X-Part: High Fidelity And Structure Coherent Shape Decomposition And CompletionXinhao Yan, Jiachen Xu, Yang Li, Changfeng Ma et al.CVPR 2026
- PartGen: Part-level 3D Generation and Reconstruction with Multi-view Diffusion ModelsMinghao Chen, Roman Shapovalov, Iro Laina, Tom Monnier et al.CVPR 2025
- NaTex: Seamless Texture Generation as Latent Color DiffusionZeqiang Lai, Yunfei Zhao, Zibo Zhao, Xin Yang et al.CVPR 2026 · 11 citations
