BodyGAN: General-purpose Controllable Neural Human Body Generation
Chaojie Yang, Hanhui Li, Shengjie Wu, Shengkai Zhang, Haonan Yan, Nianhong Jiao, Jie Tang, Runnan Zhou, Xiaodan Liang, Tianxiang Zheng
摘要
Recent advances in generative adversarial networks (GANs) have provided potential solutions for photo-realistic human image synthesis. However, the explicit and individual control of synthesis over multiple factors, such as poses, body shapes, and skin colors, remains difficult for existing methods. This is because current methods mainly rely on a single pose/appearance model, which is limited in dis-entangling various poses and appearance in human images. In addition, such a unimodal strategy is prone to causing severe artifacts in the generated images like color distortions and unrealistic textures. To tackle these issues, this paper proposes a multi-factor conditioned method dubbed BodyGAN. Specifically, given a source image, our Body-GAN aims at capturing the characteristics of the human body from multiple aspects: (i) A pose encoding branch consisting of three hybrid subnetworks is adopted, to generate the semantic segmentation based representation, the 3D surface based representation, and the key point based rep-resentation of the human body, respectively. (ii) Based on the segmentation results, an appearance encoding branch is used to obtain the appearance information of the human body parts. (iii) The outputs of these two branches are represented by user-editable condition maps, which are then processed by a generator to predict the synthesized image. In this way, our BodyGAN can achieve the fine-grained dis-entanglement of pose, body shape, and appearance, and consequently enable the explicit and effective control of syn-thesis with diverse conditions. Extensive experiments on multiple datasets and a comprehensive user study show that our BodyGAN achieves the state-of-the-art performance.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper8
- Everybody Dance NowCaroline Chan, Shiry Ginosar, Tinghui Zhou, Alexei A. EfrosICCV 2019 · 被引用 840 次
- Learning Realistic Human Reposing using Cyclic Self-Supervision with 3D Shape, Pose, and Appearance ConsistencySoubhik Sanyal, Betty J. Mohler, Alex Vorobiov, Larry Davis 等ICCV 2021 · 被引用 20 次
- Towards Photo-Realistic Virtual Try-On by Adaptively Generating↔Preserving Image ContentHan Yang, Ruimao Zhang, Xiaobao Guo, Wei Liu 等CVPR 2020
- High-Fidelity Neural Human Motion Transfer From Monocular VideoMoritz Kappel, Vladislav Golyanik, Mohamed A. Elgharib, Jann-Ole Henningson 等CVPR 2021
- Cross-Domain Correspondence Learning for Exemplar-Based Image TranslationPan Zhang, Bo Zhang, Dong Chen, Lu Yuan 等CVPR 2020
相关 Paper
- 3DHumanGAN: 3D-Aware Human Image Generation with 3D Pose MappingZhuoqian Yang, Shikai Li, Wayne Wu, Bo DaiICCV 2023 · 被引用 19 次
- XAGen: 3D Expressive Human Avatars GenerationZhongcong Xu, Jianfeng Zhang, Jun Hao Liew, Jiashi Feng 等NeurIPS 2023 · 被引用 24 次
- Controllable Person Image Synthesis With Attribute-Decomposed GANYifang Men, Yiming Mao, Yuning Jiang, Wei-Ying Ma 等CVPR 2020
- InsetGAN for Full-Body Image GenerationAnna Frühstück, Krishna Kumar Singh, Eli Shechtman, Niloy J. Mitra 等CVPR 2022 · 被引用 52 次
- 3D-Aware Generative Model for Improved Side-View Image SynthesisKyungmin Jo, Wonjoon Jin, Jaegul Choo, Hyunjoon Lee 等ICCV 2023 · 被引用 5 次
