CtrlAvatar: Controllable Avatars Generation via Disentangled Invertible Networks
Wenfeng Song, Yang Ding, Fei Hou, Shuai Li, Aimin Hao, Xia Hou
Abstract
As virtual experiences grow in popularity, the demand for realistic, personalized, and animatable human avatars increases. Traditional methods, relying on fixed templates, often produce costly avatars that lack expressiveness and realism. To overcome these challenges, we introduce Controllable Avatars generation via disentangled invertible networks (CtrlAvatar), a real-time framework for generating lifelike and customizable avatars. CtrlAvatar uses disentangled invertible networks to separate the deformation process into implicit body geometry and explicit texture components. This approach eliminates the need for repeated occupancy reconstruction, enabling detailed and coherent animations. The body geometry component ensures anatomical accuracy, while the texture component allows for complex, artifact-free clothing customization. This architecture ensures smooth integration between body movements and surface details. By optimizing transformations with position-varying offsets from the avatar’s initial Linear Blend Skinning vertices, CtrlAvatar achieves flexible, natural deformations that adapt to various scenarios. Extensive experiments show that CtrlAvatar outperforms other methods in quality, diversity, controllability, and cost-efficiency, marking a significant advancement in avatar generation.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 883d371c-e26e-403b-a83c-d2bf8a67a2d7Cited by top-tier papers2
- MonoCloth: Reconstruction and Animation of Cloth-Decoupled Human Avatars from Monocular VideosDaisheng Jin, Ying HeAAAI 2026 · 1 citation
- DECON: Reconstruction of Clothed-Geometric Multiple Humans from a Single Image via Geometry-Guided DecouplingYiming Jiang, Wenfeng Song, Shuai Li, Aimin HaoAAAI 2026
Builds on23
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 5,687 citations
- Deep Marching Tetrahedra: a Hybrid Representation for High-Resolution 3D Shape SynthesisTianchang Shen, Jun Gao, Kangxue Yin, Ming-Yu Liu et al.NeurIPS 2021 · 652 citations
- HumanNeRF: Free-viewpoint Rendering of Moving People from Monocular VideoChung-Yi Weng, Brian Curless, Pratul P. Srinivasan, Jonathan T. Barron et al.CVPR 2022 · 411 citations
- Occupancy Flow: 4D Reconstruction by Learning Particle DynamicsMichael Niemeyer, Lars M. Mescheder, Michael Oechsle, Andreas GeigerICCV 2019 · 314 citations
- SNARF: Differentiable Forward Skinning for Animating Non-Rigid Neural Implicit ShapesXu Chen, Yufeng Zheng, Michael J. Black, Otmar Hilliges et al.ICCV 2021 · 267 citations
Related papers
- Disentangled Clothed Avatar Generation with Layered RepresentationWeitian Zhang, Yichao Yan, Sijing Wu, Manwen Liao et al.ICCV 2025 · 3 citations
- GETAvatar: Generative Textured Meshes for Animatable Human AvatarsXuanmeng Zhang, Jianfeng Zhang, Rohan Chacko, Hongyi Xu et al.ICCV 2023 · 32 citations
- FastAnimate: Towards Learnable Template Construction and Pose Deformation for Fast 3D Human Avatar AnimationJian Shu, Nanjie Yao, Gangjian Zhang, Junlong Ren et al.AAAI 2026 · 1 citation
- EVA: Expressive Virtual Avatars from Multi-view VideosHendrik Junkawitsch, Guoxing Sun, Heming Zhu, Christian Theobalt et al.SIGGRAPH 2025 · 4 citations
- Relightable and Animatable Neural Avatars from VideosWenbin Lin, Chengwei Zheng, Jun-Hai Yong, Feng XuAAAI 2024 · 22 citations
