Driving-signal aware full-body avatars
Timur M. Bagautdinov, Chenglei Wu, Tomas Simon, Fabián Prada, Takaaki Shiratori, Shih-En Wei, Weipeng Xu, Yaser Sheikh, Jason M. Saragih
摘要
We present a learning-based method for building driving-signal aware full-body avatars. Our model is a conditional variational autoencoder that can be animated with incomplete driving signals, such as human pose and facial keypoints, and produces a high-quality representation of human geometry and view-dependent appearance. The core intuition behind our method is that better drivability and generalization can be achieved by disentangling the driving signals and remaining generative factors, which are not available during animation. To this end, we explicitly account for information deficiency in the driving signal by introducing a latent space that exclusively captures the remaining information, thus enabling the imputation of the missing factors required during full-body animation, while remaining faithful to the driving signal. We also propose a learnable localized compression for the driving signal which promotes better generalization, and helps minimize the influence of global chance-correlations often found in real datasets. For a given driving signal, the resulting variational model produces a compact space of uncertainty for missing factors that allows for an imputation strategy best suited to a particular application. We demonstrate the efficacy of our approach on the challenging problem of full-body animation for virtual telepresence with driving signals acquired from minimal sensors placed in the environment and mounted on a VR-headset.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper51
- Structured Local Radiance Fields for Human Avatar ModelingZerong Zheng, Han Huang, Tao Yu, Hongwen Zhang 等CVPR 2022 · 被引用 115 次
- DeepMultiCap: Performance Capture of Multiple Characters Using Sparse Multiview CamerasYang Zheng, Ruizhi Shao, Yuxiang Zhang, Tao Yu 等ICCV 2021 · 被引用 112 次
- When XR and AI Meet - A Scoping Review on Extended Reality and Artificial IntelligenceTeresa Hirzle, Florian Müller, Fiona Draxler, Martin Schmitz 等CHI 2023 · 被引用 90 次
- AvatarReX: Real-time Expressive Full-body AvatarsZerong Zheng, Xiaochen Zhao, Hongwen Zhang, Boning Liu 等SIGGRAPH 2023 · 被引用 80 次
- VirtualCube: An Immersive 3D Video Communication SystemYizhong Zhang, Jiaolong Yang, Zhen Liu, Ruicheng Wang 等IEEE VR 2022 · 被引用 66 次
它引用的顶会 Paper13
- NVAE: A Deep Hierarchical Variational AutoencoderArash Vahdat, Jan KautzNeurIPS 2020 · 被引用 1,141 次
- Everybody Dance NowCaroline Chan, Shiry Ginosar, Tinghui Zhou, Alexei A. EfrosICCV 2019 · 被引用 840 次
- Liquid Warping GAN: A Unified Framework for Human Motion Imitation, Appearance Transfer and Novel View SynthesisWen Liu, Zhixin Piao, Jie Min, Wenhan Luo 等ICCV 2019 · 被引用 285 次
- MeshSDF: Differentiable Iso-Surface ExtractionEdoardo Remelli, Artem Lukoianov, Stephan R. Richter, Benoît Guillard 等NeurIPS 2020 · 被引用 186 次
- NPMs: Neural Parametric Models for 3D Deformable ShapesPablo R. Palafox, Aljaz Bozic, Justus Thies, Matthias Nießner 等ICCV 2021 · 被引用 129 次
相关 Paper
- Full-Body Motion from a Single Head-Mounted Device: Generating SMPL Poses from Partial ObservationsAndrea Dittadi, Sebastian Dziadzio, Darren Cosker, Ben Lundell 等ICCV 2021 · 被引用 78 次
- Drivable Volumetric Avatars using Texel-Aligned FeaturesEdoardo Remelli, Timur M. Bagautdinov, Shunsuke Saito, Chenglei Wu 等SIGGRAPH 2022 · 被引用 61 次
- FLAG: Flow-based 3D Avatar Generation from Sparse ObservationsSadegh Aliakbarian, Pashmina Cameron, Federica Bogo, Andrew W. Fitzgibbon 等CVPR 2022
- NeuWigs: A Neural Dynamic Model for Volumetric Hair Capture and AnimationZiyan Wang, Giljoo Nam, Tuur Stuyck, Stephen Lombardi 等CVPR 2023
- Auto-CARD: Efficient and Robust Codec Avatar Driving for Real-time Mobile TelepresenceYonggan Fu, Yuecheng Li, Chenghui Li, Jason M. Saragih 等CVPR 2023
