SKDream: Controllable Multi-view and 3D Generation with Arbitrary Skeletons
Yuanyou Xu, Zongxin Yang, Yi Yang
摘要
Controllable generation has achieved substantial progress in both 2D and 3D domains, yet current conditional generation methods still face limitations in describing detailed shape structures. Skeletons can effectively represent and describe object anatomy and pose. Unfortunately, past studies are often limited to human skeletons. In this work, we generalize skeletal conditioned generation to arbitrary structures. First, we design a reliable mesh skeletonization pipeline to generate a large-scale mesh-skeleton paired dataset. Based on the dataset, a multi-view and 3D generation pipeline is built. We propose to represent 3D skeletons by Coordinate Color Encoding as 2D conditional images. A Skeletal Correlation Module is designed to extract global skeletal features for condition injection. After multi-view images are generated, 3D assets can be obtained by incorporating a large reconstruction model, followed by a UV texture refinement stage. As a result, our method achieves instant generation of multi-view and 3D contents that are aligned with given skeletons. The proposed techniques largely improve the object-skeleton alignment and generation quality. Project page at https://skdream3d.github.io/ .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Stroke3D: Lifting 2D strokes into rigged 3D model via latent diffusion modelsRuisi Zhao, Haoren Zheng, Zongxin Yang, Hehe Fan 等ICLR 2026 · 被引用 2 次
- PoseMaster: A Unified 3D Native Framework for Stylized Pose GenerationHongyu Yan, Kunming Luo, Weiyu Li, Kaiyi Zhang 等CVPR 2026 · 被引用 1 次
- Insert Anything: Image Insertion via In-Context Editing in DiTWensong Song, Hong Jiang, Zongxing Yang, Zheqiao Cheng 等AAAI 2026 · 被引用 1 次
它引用的顶会 Paper37
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 被引用 6,759 次
相关 Paper
- Animator-Centric Skeleton Generation on Objects with Fine-Grained DetailsMingze Sun, Cheng Zeng, Jiansong Pei, Junhao Chen 等CVPR 2026 · 被引用 10 次
- ARMO: Autoregressive Rigging for Multi-Category ObjectsMingze Sun, Shiwei Mao, Keyi Chen, Yurun Chen 等ICCV 2025 · 被引用 3 次
- Muses: Designing, Composing, Generating Nonexistent Fantasy 3D Creatures without TrainingHexiao Lu, Xiaokun Sun, Zeyu Cai, Hao Guo 等CVPR 2026
- DreamDance: Animating Human Images by Enriching 3D Geometry Cues from 2D PosesYatian Pang, Bin Zhu, Bin Lin, Mingzhe Zheng 等ICCV 2025 · 被引用 2 次
- Structured 3D Latents for Scalable and Versatile 3D GenerationJianfeng Xiang, Zelong Lv, Sicheng Xu, Yu Deng 等CVPR 2025
