CLIP-Sculptor: Zero-Shot Generation of High-Fidelity and Diverse Shapes from Natural Language
Aditya Sanghi, Rao Fu, Vivian Liu, Karl D. D. Willis, Hooman Shayani, Amir Hosein Khasahmadi, Srinath Sridhar, Daniel Ritchie
2023年份
22顶会引用
摘要
§ "a baseball cap" "a skateboard" "a motor bike" "a formula one car" "a truck that looks like a limo" "a travel bag" "an office chair" "an airplane" "a round shaped lamp" "a delta wing" Figure 1. CLIP-Sculptor is a zero-shot text-to-shape generation method. Left Top/Bottom: It generates diverse shapes that reflect the semantic meaning of the text input without requiring any text data during training. Right: The method also generates high-fidelity shapes.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper22
- Instant3D: Fast Text-to-3D with Sparse-view Generation and Large Reconstruction ModelJiahao Li, Hao Tan, Kai Zhang, Zexiang Xu 等ICLR 2024 · 被引用 408 次
- Text2CAD: Generating Sequential CAD Designs from Beginner-to-Expert Level Text PromptsMohammad Sadil Khan, Sankalp Sinha, Talha Uddin Sheikh, Didier Stricker 等NeurIPS 2024 · 被引用 148 次
- PartCrafter: Structured 3D Mesh Generation via Compositional Latent Diffusion TransformersYuchen Lin, Chenguo Lin, Panwang Pan, Honglei Yan 等NeurIPS 2025 · 被引用 89 次
- Masked Audio Generation using a Single Non-Autoregressive TransformerAlon Ziv, Itai Gat, Gaël Le Lan, Tal Remez 等ICLR 2024 · 被引用 69 次
- Make-A-Shape: a Ten-Million-scale 3D Shape ModelKa-Hei Hui, Aditya Sanghi, Arianna Rampini, Kamal Rahimi Malekshan 等ICML 2024 · 被引用 29 次
它引用的顶会 Paper21
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Zero-Shot Text-to-Image GenerationAditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray 等ICML 2021 · 被引用 6,356 次
- GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion ModelsAlexander Quinn Nichol, Prafulla Dhariwal, Aditya Ramesh, Pranav Shyam 等ICML 2022 · 被引用 4,691 次
- Mind the Gap: Understanding the Modality Gap in Multi-modal Contrastive Representation LearningWeixin Liang, Yuhui Zhang, Yongchan Kwon, Serena Yeung 等NeurIPS 2022 · 被引用 834 次
相关 Paper
- CLIP-Forge: Towards Zero-Shot Text-to-Shape GenerationAditya Sanghi, Hang Chu, Joseph G. Lambourne, Ye Wang 等CVPR 2022 · 被引用 206 次
- Towards Language-Free Training for Text-to-Image GenerationYufan Zhou, Ruiyi Zhang, Changyou Chen, Chunyuan Li 等CVPR 2022 · 被引用 182 次
- TAPS3D: Text-Guided 3D Textured Shape Generation from Pseudo SupervisionJiacheng Wei, Hao Wang, Jiashi Feng, Guosheng Lin 等CVPR 2023
- Variational Distribution Learning for Unsupervised Text-to-Image GenerationMinsoo Kang, Doyup Lee, Jiseob Kim, Saehoon Kim 等CVPR 2023
- CLIPTexture: Text-Driven Texture SynthesisYiren SongACM MM 2022 · 被引用 7 次
