UnitedHuman: Harnessing Multi-Source Data for High-Resolution Human Generation
Jianglin Fu, Shikai Li, Yuming Jiang, Kwan-Yee Lin, Wayne Wu, Ziwei Liu
摘要
Human generation has achieved significant progress. Nonetheless, existing methods still struggle to synthesize specific regions such as faces and hands. We argue that the main reason is rooted in the training data. A holistic human dataset inevitably has insufficient and low-resolution information on local parts. Therefore, we propose to use multisource datasets with various resolution images to jointly learn a high-resolution human generative model. However, multi-source data inherently a) contains different parts that do not spatially align into a coherent human, and b) comes with different scales. To tackle these challenges, we propose an end-to-end framework, UnitedHuman, that empowers continuous GAN with the ability to effectively utilize multi-source data for high-resolution human generation. Specifically, 1) we design a Multi-Source Spatial Transformer that spatially aligns multi-source images to full-body space with a human parametric model. 2) Next, a continuous GAN is proposed with global-structural guidance and CutMix consistency. Patches from different datasets are then sampled and transformed to supervise the training of this scale-invariant generative model. Extensive experiments demonstrate that our model jointly learned from multi-source data achieves superior quality than those learned from a holistic dataset. Project page: https://unitedhuman.github.io/.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Expressive Gaussian Human Avatars from Monocular RGB VideoHezhen Hu, Zhiwen Fan, Tianhao Wu, Yihan Xi 等NeurIPS 2024 · 被引用 29 次
- GauHuman: Articulated Gaussian Splatting from Monocular Human VideosShoukang Hu, Tao Hu, Ziwei LiuCVPR 2024
- VBench: Comprehensive Benchmark Suite for Video Generative ModelsZiqi Huang, Yinan He, Jiashuo Yu, Fan Zhang 等CVPR 2024
- Total Selfie: Generating Full-Body SelfiesBowei Chen, Brian Curless, Ira Kemelmacher-Shlizerman, Steven M. SeitzCVPR 2024
它引用的顶会 Paper16
- Alias-Free Generative Adversarial NetworksTero Karras, Miika Aittala, Samuli Laine, Erik Härkönen 等NeurIPS 2021 · 被引用 2,126 次
- PARE: Part Attention Regressor for 3D Human Body EstimationMuhammed Kocabas, Chun-Hao P. Huang, Otmar Hilliges, Michael J. BlackICCV 2021 · 被引用 509 次
- FreiHAND: A Dataset for Markerless Capture of Hand Pose and Shape From Single RGB ImagesChristian Zimmermann, Duygu Ceylan, Jimei Yang, Bryan C. Russell 等ICCV 2019 · 被引用 493 次
- Text2Human: text-driven controllable human image generationYuming Jiang, Shuai Yang, Haonan Qiu, Wayne Wu 等SIGGRAPH 2022 · 被引用 140 次
- Exploring Dual-task Correlation for Pose Guided Person Image GenerationPengze Zhang, Lingxiao Yang, Jianhuang Lai, Xiaohua XieCVPR 2022 · 被引用 92 次
相关 Paper
- InsetGAN for Full-Body Image GenerationAnna Frühstück, Krishna Kumar Singh, Eli Shechtman, Niloy J. Mitra 等CVPR 2022 · 被引用 52 次
- HumanCrafter: Synergizing Generalizable Human Reconstruction and Semantic 3D SegmentationPanwang Pan, Tingting Shen, Chenxin Li, Yunlong Lin 等NeurIPS 2025
- High-fidelity 3D Human Digitization from Single 2K Resolution ImagesSang-Hun Han, Min-Gyu Park, Ju Hong Yoon, Ju-Mi Kang 等CVPR 2023
- 3DHumanGAN: 3D-Aware Human Image Generation with 3D Pose MappingZhuoqian Yang, Shikai Li, Wayne Wu, Bo DaiICCV 2023 · 被引用 19 次
- BodyGAN: General-purpose Controllable Neural Human Body GenerationChaojie Yang, Hanhui Li, Shengjie Wu, Shengkai Zhang 等CVPR 2022 · 被引用 8 次
