EdgeRunner: Auto-regressive Auto-encoder for Artistic Mesh Generation
Jiaxiang Tang, Zhaoshuo Li, Zekun Hao, Xian Liu, Gang Zeng, Ming-Yu Liu, Qinsheng Zhang
Abstract
Current auto-regressive mesh generation methods suffer from issues such as incompleteness, insufficient detail, and poor generalization. In this paper, we propose an Auto-regressive Auto-encoder (ArAE) model capable of generating high-quality 3D meshes with up to 4,000 faces at a spatial resolution of . We introduce a novel mesh tokenization algorithm that efficiently compresses triangular meshes into 1D token sequences, significantly enhancing training efficiency. Furthermore, our model compresses variable-length triangular meshes into a fixed-length latent space, enabling training latent diffusion models for better generalization. Extensive experiments demonstrate the superior quality, diversity, and generalization capabilities of our model in both point cloud and image-conditioned mesh generation tasks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 30e41789-796e-4f62-b73f-5d1a9a2d7cfaCited by top-tier papers48
- PartCrafter: Structured 3D Mesh Generation via Compositional Latent Diffusion TransformersYuchen Lin, Chenguo Lin, Panwang Pan, Honglei Yan et al.NeurIPS 2025 · 89 citations
- Efficient Part-level 3D Object Generation via Dual Volume PackingJiaxiang Tang, Ruijie Lu, Max Li, Zekun Hao et al.NeurIPS 2025 · 53 citations
- Puppeteer: Rig and Animate Your 3D ModelsChaoyue Song, Xiu Li, Fan Yang, Zhongcong Xu et al.NeurIPS 2025 · 48 citations
- ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and UnderstandingJunliang Ye, Zhengyi Wang, Ruowen Zhao, Shenghao Xie et al.NeurIPS 2025 · 42 citations
- Mesh-RFT: Enhancing Mesh Generation via Fine-grained Reinforcement Fine-TuningJian Liu, Jing Xu, Song Guo, Jing Li et al.NeurIPS 2025 · 22 citations
Builds on53
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
- FlashAttention-2: Faster Attention with Better Parallelism and Work PartitioningTri DaoICLR 2024 · 2,600 citations
Related papers
- FACE: A Face-based Autoregressive Representation for High-Fidelity and Efficient Mesh GenerationHanxiao Wang, Yuanchen Guo, Ying-Tian Liu, Zi-Xin Zou et al.CVPR 2026 · 6 citations
- MeshFlow: Efficient Artistic Mesh Generation via MeshVAE and Flow-based Diffusion TransformerWeiyu Li, Antoine Toisoul, Tom Monnier, Roman Shapovalov et al.CVPR 2026 · 7 citations
- TreeMeshGPT: Artistic Mesh Generation with Autoregressive Tree SequencingStefan Lionar, Jiabin Liang, Gim Hee LeeCVPR 2025
- HiFi-Mesh: High-Fidelity Efficient 3D Mesh Generation via Compact Autoregressive DependenceYanfeng Li, Tao Tan, Qinquan Gao, Zhiwen Cao et al.AAAI 2026
- Topology-Preserved Auto-regressive Mesh Generation in the Manner of Weaving SilkGaochao Song, Zibo Zhao, Haohan Weng, Jingbo Zeng et al.ICLR 2026 · 6 citations
