Learning Motion-Dependent Appearance for High-Fidelity Rendering of Dynamic Humans from a Single Camera
Jae Shin Yoon, Duygu Ceylan, Tuanfeng Y. Wang, Jingwan Lu, Jimei Yang, Zhixin Shu, Hyun Soo Park
Abstract
Appearance of dressed humans undergoes a complex geometric transformation induced not only by the static pose but also by its dynamics, i.e., there exists a number of cloth geometric configurations given a pose depending on the way it has moved. Such appearance modeling conditioned on motion has been largely neglected in existing human rendering methods, resulting in rendering of physically implausible motion. A key challenge of learning the dynamics of the appearance lies in the requirement of a prohibitively large amount of observations. In this paper, we present a compact motion representation by enforcing equivariance—a representation is expected to be transformed in the way that the pose is transformed. We model an equivariant encoder that can generate the generalizable representation from the spatial and temporal derivatives of the 3D body surface. This learned representation is decoded by a compositional multi-task decoder that renders high fidelity time-varying appearance. Our experiments show that our method can generate a temporally coherent video of dynamic humans for unseen body poses and novel views given a single view video.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers9
- DELIFFAS: Deformable Light Fields for Fast Avatar SynthesisYoungjoong Kwon, Lingjie Liu, Henry Fuchs, Marc Habermann et al.NeurIPS 2023 · 45 citations
- PoseVocab: Learning Joint-structured Pose Embeddings for Human Avatar ModelingZhe Li, Zerong Zheng, Yuxiao Liu, Boyao Zhou et al.SIGGRAPH 2023 · 34 citations
- RANA: Relightable Articulated Neural AvatarsUmar Iqbal, Akin Caliskan, Koki Nagano, Sameh Khamis et al.ICCV 2023 · 14 citations
- Bidirectional Temporal Diffusion Model for Temporally Consistent Human AnimationTserendorj Adiya, Jae Shin Yoon, Jungeun Lee, Sanghun Kim et al.ICLR 2024 · 2 citations
- HumanRAM: Feed-forward Human Reconstruction and Animation Model using TransformersZhiyuan Yu, Zhe Li, Hujun Bao, Can Yang et al.SIGGRAPH 2025 · 2 citations
Builds on21
- Learning to Reconstruct 3D Human Pose and Shape via Model-Fitting in the LoopNikos Kolotouros, Georgios Pavlakos, Michael J. Black, Kostas DaniilidisICCV 2019 · 1,139 citations
- Everybody Dance NowCaroline Chan, Shiry Ginosar, Tinghui Zhou, Alexei A. EfrosICCV 2019 · 840 citations
- Non-Rigid Neural Radiance Fields: Reconstruction and Novel View Synthesis of a Dynamic Scene From Monocular VideoEdgar Tretschk, Ayush Tewari, Vladislav Golyanik, Michael Zollhöfer et al.ICCV 2021 · 617 citations
- Dynamic View Synthesis from Dynamic Monocular VideoChen Gao, Ayush Saraf, Johannes Kopf, Jia-Bin HuangICCV 2021 · 522 citations
- ClothFlow: A Flow-Based Model for Clothed Person GenerationXintong Han, Weilin Huang, Xiaojun Hu, Matthew R. ScottICCV 2019 · 297 citations
Related papers
- Structured Local Radiance Fields for Human Avatar ModelingZerong Zheng, Han Huang, Tao Yu, Hongwen Zhang et al.CVPR 2022 · 115 citations
- GaussianAvatar: Towards Realistic Human Avatar Modeling from a Single Video via Animatable 3D GaussiansLiangxiao Hu, Hongwen Zhang, Yuxiang Zhang, Boyao Zhou et al.CVPR 2024
- Link to the Past: Temporal Propagation for Fast 3D Human Reconstruction from Monocular VideoMatthew Marchellus, Nadhira Noor, In Kyu ParkCVPR 2025
- Animatable Virtual Humans: Learning Pose-Dependent Human Representations in UV Space for Interactive Performance SynthesisWieland Morgenstern, Milena T. Bagdasarian, Anna Hilsmann, Peter EisertIEEE VR 2024 · 7 citations
- Behavior-Driven Synthesis of Human DynamicsAndreas Blattmann, Timo Milbich, Michael Dorkenwald, Björn OmmerCVPR 2021
