HumanNorm: Learning Normal Diffusion Model for High-quality and Realistic 3D Human Generation
Xin Huang, Ruizhi Shao, Qi Zhang, Hongwen Zhang, Ying Feng, Yebin Liu, Qing Wang
摘要
Recent text-to-3D methods employing diffusion models have made significant advancements in 3D human generation. However, these approaches face challenges due to the limitations of text-to-image diffusion models, which lack an understanding of 3D structures. Consequently, these methods struggle to achieve high-quality human generation, resulting in smooth geometry and cartoon-like appearances. In this paper, we propose HumanNorm, a novel approach for high-quality and realistic 3D human generation. The main idea is to enhance the model's 2D perception of 3D geometry by learning a normal-adapted diffusion model and a normal-aligned diffusion model. The normal-adapted diffusion model can generate high-fidelity normal maps corresponding to user prompts with view- † Work done during an internship at Tsinghua University.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper29
- HumanSplat: Generalizable Single-Image Human Gaussian Splatting with Structure PriorsPanwang Pan, Zhuo Su, Chenguo Lin, Zhen Fan 等NeurIPS 2024 · 被引用 76 次
- DreamMat: High-quality PBR Material Generation with Geometry- and Light-aware Diffusion ModelsYuqing Zhang, Yuan Liu, Zhiyu Xie, Lei Yang 等SIGGRAPH 2024 · 被引用 28 次
- LHM: Large Animatable Human Reconstruction Model for Single Image to 3D in SecondsLingteng Qiu, Xiaodong Gu, Peihao Li, Qi Zuo 等ICCV 2025 · 被引用 19 次
- Control4D: Efficient 4D Portrait Editing With TextRuizhi Shao, Jingxiang Sun, Cheng Peng, Zerong Zheng 等CVPR 2024 · 被引用 17 次
- Hi3dgen: High-Fidelity 3D Geometry Generation From Images Via Normal BridgingChongjie Ye, Yushuang Wu, Ziteng Lu, Jiahao Chang 等ICCV 2025 · 被引用 11 次
它引用的顶会 Paper37
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 被引用 6,759 次
- Zero-Shot Text-to-Image GenerationAditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray 等ICML 2021 · 被引用 6,356 次
相关 Paper
- HumanRef: Single Image to 3D Human Generation via Reference-Guided DiffusionJingbo Zhang, Xiaoyu Li, Qi Zhang, Yanpei Cao 等CVPR 2024 · 被引用 15 次
- HumanPro: Single-view 3D Clothed Human Reconstruction with Progressive Normal GuidanceJianchi Sun, Fei Luo, Wenzhuo Fan, Yu Jiang 等AAAI 2026
- GeneMAN: Generalizable Single-Image 3D Human Reconstruction from Multi-Source Human DataWentao Wang, Hang Ye, Fangzhou Hong, Xue Yang 等NeurIPS 2025 · 被引用 6 次
- Chupa: Carving 3D Clothed Humans from Skinned Shape Priors using 2D Diffusion Probabilistic ModelsByungjun Kim, Patrick Kwon, Kwangho Lee, Myunggi Lee 等ICCV 2023 · 被引用 26 次
- EfficientDreamer: High-Fidelity and Stable 3D Creation via Orthogonal-view Diffusion PriorsZhipeng Hu, Minda Zhao, Chaoyi Zhao, Xinyue Liang 等CVPR 2024
