PrimitiveAnything: Human-Crafted 3D Primitive Assembly Generation with Auto-Regressive transformer
Jingwen Ye, Yuze He, Yanning Zhou, Yiqin Zhu, Kaiwen Xiao, Yong-Jin Liu, Wei Yang, Xiao Han
Abstract
Shape primitive abstraction, which decomposes complex 3D shapes into simple geometric elements, plays a crucial role in human visual cognition and has broad applications in computer vision and graphics. While recent advances in 3D content generation have shown remarkable progress, existing primitive abstraction methods either rely on geometric optimization with limited semantic understanding or learn from small-scale, category-specific datasets, struggling to generalize across diverse shape categories. We present PrimitiveAnything, a novel framework that reformulates shape primitive abstraction as a primitive assembly generation task. PrimitiveAnything includes a shape-conditioned primitive transformer for auto-regressive generation and an ambiguity-free parameterization scheme to represent multiple types of primitives in a unified manner. The proposed framework directly learns the process of primitive assembly from large-scale human-crafted abstractions, enabling it to capture how humans decompose complex shapes into primitive elements. Through extensive experiments, we demonstrate that PrimitiveAnything can generate high-quality primitive assemblies that better align with human perception while maintaining geometric fidelity across diverse shape categories. It benefits various 3D applications and shows potential for enabling primitive-based user-generated content (UGC) in games. Project page: https://primitiveanything.github.io
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 691d897a-b75c-4f4e-9636-d6c312053474Cited by top-tier papers6
- PartCrafter: Structured 3D Mesh Generation via Compositional Latent Diffusion TransformersYuchen Lin, Chenguo Lin, Panwang Pan, Honglei Yan et al.NeurIPS 2025 · 89 citations
- SPARK: Sim-ready Part-level Articulated Reconstruction with VLM KnowledgeYumeng He, Ying Jiang, Jiayin Lu, Yin Yang et al.CVPR 2026 · 6 citations
- Residual Primitive Fitting of 3D Shapes with SuperFrustaAditya Ganeshan, Matheus Gadelha, Thibault Groueix, Zhiqin Chen et al.CVPR 2026 · 5 citations
- Prox-E: Fine-Grained 3D Shape Editing via Primitive-Based AbstractionsEtai Sella, Hao Phung, Nitay Amiel, Or Litany et al.SIGGRAPH 2026 · 2 citations
- AssetFormer: Modular 3D Assets Generation with Autoregressive TransformerLingting Zhu, Shengju Qian, Haidi Fan, Jiayu Dong et al.ICLR 2026 · 1 citation
Builds on37
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Scalable Diffusion Models with TransformersWilliam Peebles, Saining XieICCV 2023 · 5,568 citations
- SDXL: Improving Latent Diffusion Models for High-Resolution Image SynthesisDustin Podell, Zion English, Kyle Lacey, Andreas Blattmann et al.ICLR 2024 · 4,569 citations
- ProlificDreamer: High-Fidelity and Diverse Text-to-3D Generation with Variational Score DistillationZhengyi Wang, Cheng Lu, Yikai Wang, Fan Bao et al.NeurIPS 2023 · 1,498 citations
- Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale PredictionKeyu Tian, Yi Jiang, Zehuan Yuan, Bingyue Peng et al.NeurIPS 2024 · 1,199 citations
Related papers
- DeFormer: Integrating Transformers with Deformable Models for 3D Shape Abstraction from a Single ImageDi Liu, Xiang Yu, Meng Ye, Qilong Zhangli et al.ICCV 2023 · 14 citations
- Unsupervised learning for cuboid shape abstraction via joint segmentation from point cloudsKaizhi Yang, Xuejin ChenSIGGRAPH 2021 · 52 citations
- ShapeCoder: Discovering Abstractions for Visual Programs from Unstructured PrimitivesR. Kenny Jones, Paul Guerrero, Niloy J. Mitra, Daniel RitchieSIGGRAPH 2023 · 19 citations
- Neural Parts: Learning Expressive 3D Shape Abstractions With Invertible Neural NetworksDespoina Paschalidou, Angelos Katharopoulos, Andreas Geiger, Sanja FidlerCVPR 2021
- RigAnything: Template-Free Autoregressive Rigging for Diverse 3D AssetsIsabella Liu, Zhan Xu, Wang Yifan, Hao Tan et al.SIGGRAPH 2025 · 11 citations
