From One to More: Contextual Part Latents for 3D Generation
Shaocong Dong, Lihe Ding, Xiao Chen, Yaokun Li, Yuxin Wang, Yucheng Wang, Qi Wang, Jaehyeok Kim, Chenjian Gao, Zhanpeng Huang, Zibin Wang, Tianfan Xue, Dan Xu
摘要
To generate 3D objects, early research focused on multi-view-driven approaches relying solely on 2D renderings. Recently, the 3D native latent diffusion paradigm has demonstrated superior performance in 3D generation, because it fully leverages the geometric information provided in ground truth 3D data. Despite its fast development, 3D diffusion still faces three challenges. First, the majority of these methods represent a 3D object by one single latent, regardless of its complexity. This may lead to detail loss when generating 3D objects with multiple complicated parts. Second, most 3D assets are designed parts by parts, yet the current holistic latent representation overlooks the independence of these parts and their interrelationships, limiting the model's generative ability. Third, current methods rely on global conditions (e.g., text, image, point cloud) to control the generation process, lacking detailed controllability. Therefore, motivated by how 3D designers create a 3D object, we present a new part-based 3D generation framework, CoPart, which represents a 3D object with multiple contextual part latents and simultaneously generates coherent 3D parts. This part-based framework has several advantages, including: i) reduces the encoding burden of intricate objects by decomposing them into simpler parts, ii) facilitates part learning and part relationship modeling, and iii) naturally supports part-level control. Furthermore, to ensure the coherence of part latents and to harness the powerful priors from foundation models, we propose a novel mutual guidance strategy to fine-tune pre-trained diffusion models for joint part latent denoising. Benefiting from the part-based representation, we demonstrate that CoPart can support various applications including part-editing, articulated object generation, and mini-scene generation. Moreover, we collect a new large-scale 3D part dataset named Partverse from Objaverse through automatic mesh segmentation and subsequent human post-annotations. By training on the proposed dataset, CoPart achieves promising part-based 3D generation with high controllability. Project page: https://hkdsc.github.io/project/copart.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- FullPart: Generating each 3D Part at Full ResolutionLihe Ding, Shaocong Dong, Yaokun Li, Chenjian Gao 等ICLR 2026 · 被引用 17 次
- PhysForge: Generating Physics-Grounded 3D Assets for Interactive Virtual WorldYunhan Yang, Chunshi Wang, Junliang Ye, YANG LI 等ICML 2026 · 被引用 6 次
- M3DLayout: A Multi-Source Dataset of 3D Indoor Layouts and Structured Descriptions for 3D GenerationYiheng Zhang, Zhuojiang Cai, Mingdao Wang, Meitong Guo 等CVPR 2026 · 被引用 5 次
- Repurposing 3D Generative Model for Autoregressive Layout GenerationHaoran Feng, Yifan Niu, Zehuan Huang, Yangtian Sun 等CVPR 2026 · 被引用 3 次
- JRM: Joint Reconstruction Model for Multiple Objects without AlignmentQirui Wu, Mohd Yawar Nihal Siddiqui, Duncan Frost, Samir Aroudj 等CVPR 2026 · 被引用 2 次
它引用的顶会 Paper34
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao 等ICCV 2023 · 被引用 13,211 次
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 被引用 6,759 次
- Scalable Diffusion Models with TransformersWilliam Peebles, Saining XieICCV 2023 · 被引用 5,568 次
- Zero-1-to-3: Zero-shot One Image to 3D ObjectRuoshi Liu, Rundi Wu, Basile Van Hoorick, Pavel Tokmakov 等ICCV 2023 · 被引用 1,662 次
相关 Paper
- UniPart: Part-Level 3D Generation with Unified 3D Geom-Seg LatentsXufan He, Yushuang Wu, Xiaoyang Guo, Chongjie Ye 等CVPR 2026 · 被引用 9 次
- HoloPart: Generative 3D Part Amodal SegmentationYunhan Yang, Yuanchen Guo, Yukun Huang, Zi-Xin Zou 等ICLR 2026 · 被引用 62 次
- X-Part: High Fidelity And Structure Coherent Shape Decomposition And CompletionXinhao Yan, Jiachen Xu, Yang Li, Changfeng Ma 等CVPR 2026
- PartGen: Part-level 3D Generation and Reconstruction with Multi-view Diffusion ModelsMinghao Chen, Roman Shapovalov, Iro Laina, Tom Monnier 等CVPR 2025
- NaTex: Seamless Texture Generation as Latent Color DiffusionZeqiang Lai, Yunfei Zhao, Zibo Zhao, Xin Yang 等CVPR 2026 · 被引用 11 次
