MonoCloth: Reconstruction and Animation of Cloth-Decoupled Human Avatars from Monocular Videos
Daisheng Jin, Ying He
Abstract
Reconstructing realistic 3D human avatars from monocular videos is a challenging task due to the limited geometric information and complex non-rigid motion involved. We present MonoCloth, a new method for reconstructing and animating clothed human avatars from monocular videos. To overcome the limitations of monocular input, we introduce a part-based decomposition strategy that separates the avatar into body, face, hands, and clothing. This design reflects the varying levels of reconstruction difficulty and deformation complexity across these components. Specifically, we focus on detailed geometry recovery for the face and hands. For clothing, we propose a dedicated cloth simulation module that captures garment deformation using temporal motion cues and geometric constraints. Experimental results demonstrate that MonoCloth improves both visual reconstruction quality and animation realism compared to existing methods. Furthermore, thanks to its part-based design, MonoCloth also supports additional tasks such as clothing transfer, underscoring its versatility and practical utility.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 28a7679e-022c-4a76-8847-1fe09eea7394Builds on18
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 5,687 citations
- Animatable Neural Radiance Fields for Modeling Dynamic Human BodiesSida Peng, Junting Dong, Qianqian Wang, Shangzhan Zhang et al.ICCV 2021 · 461 citations
- HumanNeRF: Free-viewpoint Rendering of Moving People from Monocular VideoChung-Yi Weng, Brian Curless, Pratul P. Srinivasan, Jonathan T. Barron et al.CVPR 2022 · 411 citations
- 3DGS-Avatar: Animatable Avatars via Deformable 3D Gaussian SplattingZhiyin Qian, Shaofei Wang, Marko Mihajlovic, Andreas Geiger et al.CVPR 2024 · 131 citations
- AvatarReX: Real-time Expressive Full-body AvatarsZerong Zheng, Xiaochen Zhao, Hongwen Zhang, Boning Liu et al.SIGGRAPH 2023 · 80 citations
Related papers
- D^3-Human: Dynamic Disentangled Digital Human from Monocular VideoHonghu Chen, Bo Peng, Yunfan Tao, Juyong ZhangCVPR 2025
- IntrinsicAvatar: Physically Based Inverse Rendering of Dynamic Humans from Monocular Videos via Explicit Ray TracingShaofei Wang, Bozidar Antic, Andreas Geiger, Siyu TangCVPR 2024 · 13 citations
- DLCA-Recon: Dynamic Loose Clothing Avatar Reconstruction from Monocular VideosChunjie Luo, Fei Luo, Yusen Wang, Enxu Zhao et al.AAAI 2024 · 5 citations
- DeClotH: Decomposable 3D Cloth and Human Body Reconstruction from a Single ImageHyeongjin Nam, Donghwan Kim, Jeongtaek Oh, Kyoung Mu LeeCVPR 2025
- Vid2Avatar: 3D Avatar Reconstruction from Videos in the Wild via Self-supervised Scene DecompositionChen Guo, Tianjian Jiang, Xu Chen, Jie Song et al.CVPR 2023
