Towards Implicit Text-Guided 3D Shape Generation
Zhengzhe Liu, Yi Wang, Xiaojuan Qi, Chi-Wing Fu
摘要
In this work, we explore the challenging task of generating 3D shapes from text. Beyond the existing works, we propose a new approach for text-guided 3D shape generation, capable of producing high-fidelity shapes with colors that match the given text description. This work has several technical contributions. First, we decouple the shape and color predictions for learning features in both texts and shapes, and propose the word-level spatial transformer to correlate word features from text with spatial features from shape. Also, we design a cyclic loss to encourage consistency between text and shape, and introduce the shape IMLE to diversify the generated shapes. Further, we extend the framework to enable text-guided shape manipulation. Extensive experiments on the largest existing text-shape benchmark [11] manifest the superiority of this work. The code and the models are available at https://github.com/liuzhengzhe/Towards-Implicit-Text-Guided-Shape-Generation .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper33
- DreamFusion: Text-to-3D using 2D DiffusionBen Poole, Ajay Jain, Jonathan T. Barron, Ben MildenhallICLR 2023 · 被引用 463 次
- Instant3D: Fast Text-to-3D with Sparse-view Generation and Large Reconstruction ModelJiahao Li, Hao Tan, Kai Zhang, Zexiang Xu 等ICLR 2024 · 被引用 408 次
- CLIP2Point: Transfer CLIP to Point Cloud Classification with Image-Depth Pre-TrainingTianyu Huang, Bowen Dong, Yunhan Yang, Xiaoshui Huang 等ICCV 2023 · 被引用 220 次
- ShapeCrafter: A Recursive Text-Conditioned 3D Shape Generation ModelRao Fu, Xiao Zhan, Yiwen Chen, Daniel Ritchie 等NeurIPS 2022 · 被引用 98 次
- VPP: Efficient Conditional 3D Generation via Voxel-Point Progressive RepresentationZekun Qi, Muzhou Yu, Runpei Dong, Kaisheng MaNeurIPS 2023 · 被引用 22 次
它引用的顶会 Paper30
- StyleCLIP: Text-Driven Manipulation of StyleGAN ImageryOr Patashnik, Zongze Wu, Eli Shechtman, Daniel Cohen-Or 等ICCV 2021 · 被引用 1,437 次
- Image2StyleGAN: How to Embed Images Into the StyleGAN Latent Space?Rameen Abdal, Yipeng Qin, Peter WonkaICCV 2019 · 被引用 1,195 次
- Multiview Neural Surface Reconstruction by Disentangling Geometry and AppearanceLior Yariv, Yoni Kasten, Dror Moran, Meirav Galun 等NeurIPS 2020 · 被引用 1,010 次
- Revisiting Point Cloud Classification: A New Benchmark Dataset and Classification Model on Real-World DataMikaela Angelina Uy, Quang-Hieu Pham, Binh-Son Hua, Duc Thanh Nguyen 等ICCV 2019 · 被引用 1,003 次
- Neural Unsigned Distance Fields for Implicit Function LearningJulian Chibane, Aymen Mir, Gerard Pons-MollNeurIPS 2020 · 被引用 415 次
相关 Paper
- ISS: Image as Stepping Stone for Text-Guided 3D Shape GenerationZhengzhe Liu, Peng Dai, Ruihui Li, Xiaojuan Qi 等ICLR 2023 · 被引用 4 次
- Diffusion-SDF: Text-to-Shape via Voxelized DiffusionMuheng Li, Yueqi Duan, Jie Zhou, Jiwen LuCVPR 2023
- ShapeWords: Guiding Text-to-Image Synthesis with 3D Shape-Aware PromptsDmitry Petrov, Pradyumn Goyal, Divyansh Shivashok, Yuanming Tao 等CVPR 2025
- Michelangelo: Conditional 3D Shape Generation based on Shape-Image-Text Aligned Latent RepresentationZibo Zhao, Wen Liu, Xin Chen, Xianfang Zeng 等NeurIPS 2023 · 被引用 279 次
- ShapeScaffolder: Structure-Aware 3D Shape Generation from TextXi Tian, Yong-Liang Yang, Qi WuICCV 2023 · 被引用 14 次
