EgoRenderer: Rendering Human Avatars from Egocentric Camera Images
Tao Hu, Kripasindhu Sarkar, Lingjie Liu, Matthias Zwicker, Christian Theobalt
Abstract
We present EgoRenderer, a system for rendering fullbody neural avatars of a person captured by a wearable, egocentric fisheye camera that is mounted on a cap or a VR headset. Our system renders photorealistic novel views of the actor and her motion from arbitrary virtual camera locations. Rendering full-body avatars from such egocentric images come with unique challenges due to the topdown view and large distortions. We tackle these challenges by decomposing the rendering process into several steps, including texture synthesis, pose construction, and neural image translation. For texture synthesis, we propose Ego-DPNet, a neural network that infers dense correspondences between the input fisheye images and an underlying parametric body model, and to extract textures from egocentric inputs. In addition, to encode dynamic appearances, our approach also learns an implicit texture stack that captures detailed appearance variation across poses and viewpoints. For correct pose generation, we first estimate body pose from the egocentric view using a parametric model. We then synthesize an external free-viewpoint pose image by projecting the parametric model to the user-specified target viewpoint. We next combine the target pose image and the textures into a combined feature image, which is transformed into the output color image using a neural image translation network. Experimental evaluations show that EgoRenderer is capable of generating realistic free-viewpoint avatars of a person wearing an egocentric camera. Comparisons to several baselines demonstrate the advantages of our approach. Project page: https://vcai.mpi-inf. mpg.de/projects/EgoRenderer/ .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers4
- HDR-GS: Efficient High Dynamic Range Novel View Synthesis at 1000x Speed via Gaussian SplattingYuanhao Cai, Zihao Xiao, Yixun Liang, Minghan Qin et al.NeurIPS 2024 · 48 citations
- Robust Egocentric Photo-realistic Facial Expression Transfer for Virtual RealityAmin Jourabloo, Fernando De la Torre, Jason M. Saragih, Shih-En Wei et al.CVPR 2022 · 13 citations
- Ground Reaction Inertial Poser: Physics-based Human Motion Capture from Sparse IMUs and Insole Pressure SensorsRyosuke Hori, Jyun-Ting Song, Zhengyi Luo, Jinkun Cao et al.CVPR 2026 · 2 citations
- SurMo: Surface-based 4D Motion Modeling for Dynamic Human RenderingTao Hu, Fangzhou Hong, Ziwei LiuCVPR 2024
Builds on4
- Neural Sparse Voxel FieldsLingjie Liu, Jiatao Gu, Kyaw Zaw Lin, Tat-Seng Chua et al.NeurIPS 2020 · 1,535 citations
- Everybody Dance NowCaroline Chan, Shiry Ginosar, Tinghui Zhou, Alexei A. EfrosICCV 2019 · 840 citations
- Ego-Pose Estimation and Forecasting As Real-Time PD ControlYe Yuan, Kris KitaniICCV 2019 · 147 citations
- xR-EgoPose: Egocentric 3D Human Pose From an HMD CameraDenis Tomè, Patrick Peluse, Lourdes Agapito, Hernán BadinoICCV 2019 · 140 citations
Related papers
- High-Fidelity Human Avatars from a Single RGB CameraHao Zhao, Jinsong Zhang, Yu-Kun Lai, Zerong Zheng et al.CVPR 2022 · 38 citations
- Egocentric Whole-Body Motion Capture with FisheyeViT and Diffusion-Based Motion RefinementJian Wang, Zhe Cao, Diogo C. Luvizon, Lingjie Liu et al.CVPR 2024 · 18 citations
- Scene-Aware Egocentric 3D Human Pose EstimationJian Wang, Diogo C. Luvizon, Weipeng Xu, Lingjie Liu et al.CVPR 2023
- StylePeople: A Generative Model of Fullbody Human AvatarsArtur Grigorev, Karim Iskakov, Anastasia Ianina, Renat Bashirov et al.CVPR 2021
- HumanNeRF: Free-viewpoint Rendering of Moving People from Monocular VideoChung-Yi Weng, Brian Curless, Pratul P. Srinivasan, Jonathan T. Barron et al.CVPR 2022 · 411 citations
