Rotationally-Consistent Novel View Synthesis for Humans
Youngjoong Kwon, Stefano Petrangeli, Dahun Kim, Haoliang Wang, Henry Fuchs, Viswanathan Swaminathan
摘要
Human novel view synthesis aims to synthesize target views of a human subject given input images taken from one or more reference viewpoints. Despite significant advances in model-free novel view synthesis, existing methods present two major limitations when applied to complex shapes like humans. First, these methods mainly focus on simple and symmetric objects, e.g., cars and chairs, limiting their performances to fine-grained and asymmetric shapes. Second, existing methods cannot guarantee visual consistency across different adjacent views of the same object. To solve these problems, we present in this paper a learning framework for the novel view synthesis of human subjects, which explicitly enforces consistency across different generated views of the subject. Specifically, we introduce a novel multi-view supervision and an explicit rotational loss during the learning process, enabling the model to preserve detailed body parts and to achieve consistency between adjacent synthesized views. To show the superior performance of our approach, we present qualitative and quantitative results on the Multi-View Human Action (MVHA) dataset we collected (consisting of 3D human models animated with different Mocap sequences and captured from 54 different viewpoints), the Pose-Varying Human Model (PVHM) dataset, and ShapeNet. The qualitative and quantitative results demonstrate that our approach outperforms the state-of-the-art baselines in both per-view synthesis quality, and in preserving rotational consistency and complex shapes (e.g. fine-grained details, challenging poses) across multiple adjacent views in a variety of scenarios, for both humans and rigid objects.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper3
- Neural Human Performer: Learning Generalizable Radiance Fields for Human Performance RenderingYoungjoong Kwon, Dahun Kim, Duygu Ceylan, Henry FuchsNeurIPS 2021 · 被引用 224 次
- Neural Image-based Avatars: Generalizable Radiance Fields for Human Avatar ModelingYoungjoong Kwon, Dahun Kim, Duygu Ceylan, Henry FuchsICLR 2023 · 被引用 1 次
- Neural Body: Implicit Neural Representations With Structured Latent Codes for Novel View Synthesis of Dynamic HumansSida Peng, Yuanqing Zhang, Yinghao Xu, Qianqian Wang 等CVPR 2021
相关 Paper
- MVHumanNet: A Large-Scale Dataset of Multi-View Daily Dressing Human CapturesZhangyang Xiong, Chenghong Li, Kenkun Liu, Hongjie Liao 等CVPR 2024 · 被引用 16 次
- View-LSTM: Novel-View Video Synthesis Through View DecompositionMohamed Ilyes Lakhal, Oswald Lanz, Andrea CavallaroICCV 2019 · 被引用 13 次
- NViST: In the Wild New View Synthesis from a Single Image with TransformersWonbong Jang, Lourdes AgapitoCVPR 2024
- iVS-Net: Learning Human View Synthesis from Internet VideosJunting Dong, Qi Fang, Tianshuo Yang, Qing Shuai 等ICCV 2023 · 被引用 9 次
- Novel View Synthesis with Diffusion ModelsDaniel Watson, William Chan, Ricardo Martin-Brualla, Jonathan Ho 等ICLR 2023 · 被引用 63 次
