ARMO: Autoregressive Rigging for Multi-Category Objects
Mingze Sun, Shiwei Mao, Keyi Chen, Yurun Chen, Shunlin Lu, Jingbo Wang, Junting Dong, Ruqi Huang
Abstract
Recent advancements in large-scale generative models have significantly improved the quality and diversity of 3D shape generation. However, most existing methods focus primarily on generating static 3D models, overlooking the potentially dynamic nature of certain shapes, such as humanoids, animals, and insects. To address this gap, we focus on rigging, a fundamental task in animation that establishes skeletal structures and skinning for 3D models. In this paper, we introduce OmniRig, the first large-scale rigging dataset, comprising 79,499 meshes with detailed skeleton and skinning information. Unlike traditional benchmarks that rely on predefined standard poses (e.g., A-pose, T-pose), our dataset embraces diverse shape categories, styles, and poses. Leveraging this rich dataset, we propose ARMO, a novel rigging framework that utilizes an autoregressive model to predict both joint positions and connectivity relationships in a unified manner. By treating the skeletal structure as a complete graph and discretizing it into tokens, we encode the joints using an auto-encoder to obtain a latent embedding and an autoregressive model to predict the tokens. A mesh-conditioned latent diffusion model is used to predict the latent embedding for conditional skeleton generation. Our method addresses the limitations of regression-based approaches, which often suffer from error accumulation and suboptimal connectivity estimation. Through extensive experiments on the OmniRig dataset, our approach achieves state-of-the-art performance in skeleton prediction, demonstrating improved generalization across diverse object categories. The code and dataset will be made public for academic use upon acceptance.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8b061a20-24d1-4f5c-be28-8403499a3bc6Cited by top-tier papers1
Ask how each one uses itBuilds on28
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Wonder3D: Single Image to 3D Using Cross-Domain DiffusionXiaoxiao Long, Yuan-Chen Guo, Cheng Lin, Yuan Liu et al.CVPR 2024 · 269 citations
- Unique3D: High-Quality and Efficient 3D Mesh Generation from a Single ImageKailu Wu, Fangfu Liu, Zhihan Cai, Runjie Yan et al.NeurIPS 2024 · 185 citations
- Autoregressive Image Generation using Residual QuantizationDoyup Lee, Chiheon Kim, Saehoon Kim, Minsu Cho et al.CVPR 2022 · 184 citations
- 4DComplete: Non-Rigid Motion Estimation Beyond the Observable SurfaceYang Li, Hikari Takehara, Takafumi Taketomi, Bo Zheng et al.ICCV 2021 · 160 citations
Related papers
- Anymate: A Dataset and Baselines for Learning 3D Object RiggingYufan Deng, Yuhao Zhang, Chen Geng, Shangzhe Wu et al.SIGGRAPH 2025 · 5 citations
- One Model to Rig Them All: Diverse Skeleton Rigging with UniRigJia-Peng Zhang, Cheng-Feng Pu, Meng-Hao Guo, Yan-Pei Cao et al.SIGGRAPH 2025 · 10 citations
- HumanRig: Learning Automatic Rigging for Humanoid Character in a Large Scale DatasetZedong Chu, Feng Xiong, Meiduo Liu, Jinzhi Zhang et al.CVPR 2025
- RigAnything: Template-Free Autoregressive Rigging for Diverse 3D AssetsIsabella Liu, Zhan Xu, Wang Yifan, Hao Tan et al.SIGGRAPH 2025 · 11 citations
- RigMo: Unifying Rig and Motion Learning for Generative AnimationHao Zhang, Jiahao Luo, Bohui Wan, Yizhou Zhao et al.CVPR 2026 · 6 citations
