3DHumanEdit: Multi-modal Body Part-aware Conditioning Information Integration for 3D Human Manipulation
Feifan Xu, Tianyi Chen, Fan Yang, Yunfei Zhang, Si Wu
摘要
The rapid advancement of 3D Generative Adversarial Networks (GANs) has significantly enhanced the diversity and quality of generated 3D images. Despite these breakthroughs, the manipulation capabilities of 3D GANs remain unexplored, presenting substantial challenges for practical applications where user interaction and modification are essential. Current manipulation methods often lack the precision needed for fine-grained attribute manipulation, and struggle to maintain multi-view consistency during the editing process. To address these limitations, we propose 3DHumanEdit, a novel approach for 3D human body part-aware manipulation. 3DHumanEdit leverages multi-modal feature fusion and body part-aware feature alignment to achieve precise manipulation of individual body parts based on detailed text inputs and segmentation images. By exploring 3D prior for accurate editing and enforcing correspondence in latent space, 3DHu-manEdit ensures coherence across multiple views. Experiments demonstrate that 3DHumanEdit outperforms existing methods in both editing fidelity and multi-view consistency, offering a robust solution for fine-grained 3D manipulation.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper29
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- InstructBLIP: Towards General-purpose Vision-Language Models with Instruction TuningWenliang Dai, Junnan Li, Dongxu Li, Anthony Meng Huat Tiong 等NeurIPS 2023 · 被引用 4,013 次
- Scaling Rectified Flow Transformers for High-Resolution Image SynthesisPatrick Esser, Sumith Kulal, Andreas Blattmann, Rahim Entezari 等ICML 2024 · 被引用 3,620 次
- Training Generative Adversarial Networks with Limited DataTero Karras, Miika Aittala, Janne Hellsten, Samuli Laine 等NeurIPS 2020 · 被引用 2,345 次
- StyleCLIP: Text-Driven Manipulation of StyleGAN ImageryOr Patashnik, Zongze Wu, Eli Shechtman, Daniel Cohen-Or 等ICCV 2021 · 被引用 1,437 次
相关 Paper
- AttriHuman-3D: Editable 3D Human Avatar Generation with Attribute Decomposition and IndexingFan Yang, Tianyi Chen, Xiaosheng He, Zhongang Cai 等CVPR 2024 · 被引用 6 次
- Quantitative Manipulation of Custom Attributes on 3D-Aware Image SynthesisHoseok Do, Eunkyung Yoo, Taehyeong Kim, Chul Lee 等CVPR 2023
- Reference-Based 3D-Aware Image Editing with TriplanesBahri Batuhan Bilecen, Yigit Yalin, Ning Yu, Aysegul DundarCVPR 2025
- HumanGen: Generating Human Radiance Fields with Explicit PriorsSuyi Jiang, Haoran Jiang, Ziyu Wang, Haimin Luo 等CVPR 2023
- 3DHumanGAN: 3D-Aware Human Image Generation with 3D Pose MappingZhuoqian Yang, Shikai Li, Wayne Wu, Bo DaiICCV 2023 · 被引用 19 次
