Learning Locally Editable Virtual Humans
Hsuan-I Ho, Lixin Xue, Jie Song, Otmar Hilliges
Abstract
In this paper, we propose a novel hybrid representation and end-to-end trainable network architecture to model fully editable and customizable neural avatars. At the core of our work lies a representation that combines the modeling power of neural fields with the ease of use and inherent 3D consistency of skinned meshes. To this end, we construct a trainable feature codebook to store local geometry and texture features on the vertices of a deformable body model, thus exploiting its consistent topology under articulation. This representation is then employed in a generative auto-decoder architecture that admits fitting to unseen scans and sampling of realistic avatars with varied appearances and geometries. Furthermore, our representation allows local editing by swapping local features between 3D assets. To verify our method for avatar creation and editing, we contribute a new high-quality dataset, dubbed CustomHumans, for training and evaluation. Our experiments quantitatively and qualitatively show that our method generates diverse detailed avatars and achieves better model fitting performance compared to state-of-the-art methods. Our code and dataset are available at https: //ait.ethz.ch/custom-humans .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers34
- SplattingAvatar: Realistic Real-Time Human Avatars With Mesh-Embedded Gaussian SplattingZhijing Shao, Zhaolong Wang, Zhuang Li, Duotun Wang et al.CVPR 2024 · 92 citations
- Human-3Diffusion: Realistic Avatar Creation via Explicit 3D Consistent Diffusion ModelsYuxuan Xue, Xianghui Xie, Riccardo Marin, Gerard Pons-MollNeurIPS 2024 · 49 citations
- MagicMan: Generative Novel View Synthesis of Humans with 3D-Aware Diffusion and Iterative RefinementXu He, Zhiyong Wu, Xiaoyu Li, Di Kang et al.AAAI 2025 · 11 citations
- HAVE-FUN: Human Avatar Reconstruction from Few-Shot Unconstrained ImagesXihe Yang, Xingyu Chen, Daiheng Gao, Shaohui Wang et al.CVPR 2024 · 11 citations
- UP2You: Fast Reconstruction of Yourself from Unconstrained Photo CollectionsZeyu Cai, Ziyang Li, Xiaoben Li, Boqian Li et al.ICLR 2026 · 11 citations
Builds on36
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 4,089 citations
- Alias-Free Generative Adversarial NetworksTero Karras, Miika Aittala, Samuli Laine, Erik Härkönen et al.NeurIPS 2021 · 2,126 citations
- PIFu: Pixel-Aligned Implicit Function for High-Resolution Clothed Human DigitizationShunsuke Saito, Zeng Huang, Ryota Natsume, Shigeo Morishima et al.ICCV 2019 · 1,411 citations
- Everybody Dance NowCaroline Chan, Shiry Ginosar, Tinghui Zhou, Alexei A. EfrosICCV 2019 · 840 citations
- AI Choreographer: Music Conditioned 3D Dance Generation with AIST++Ruilong Li, Shan Yang, David A. Ross, Angjoo KanazawaICCV 2021 · 701 citations
Related papers
- X-Avatar: Expressive Human AvatarsKaiyue Shen, Chen Guo, Manuel Kaufmann, Juan Jose Zarate et al.CVPR 2023
- StylePeople: A Generative Model of Fullbody Human AvatarsArtur Grigorev, Karim Iskakov, Anastasia Ianina, Renat Bashirov et al.CVPR 2021
- Single-Shot Implicit Morphable Faces with Consistent Texture ParameterizationConnor Z. Lin, Koki Nagano, Jan Kautz, Eric R. Chan et al.SIGGRAPH 2023 · 14 citations
- CtrlAvatar: Controllable Avatars Generation via Disentangled Invertible NetworksWenfeng Song, Yang Ding, Fei Hou, Shuai Li et al.AAAI 2025 · 1 citation
- GETAvatar: Generative Textured Meshes for Animatable Human AvatarsXuanmeng Zhang, Jianfeng Zhang, Rohan Chacko, Hongyi Xu et al.ICCV 2023 · 32 citations
