3D Highlighter: Localizing Regions on 3D Shapes via Text Descriptions
Dale Decatur, Itai Lang, Rana Hanocka
Abstract
We present 3D Highlighter, a technique for localizing semantic regions on a mesh using text as input. A key feature of our system is the ability to interpret “out-of-domain” localizations. Our system demonstrates the ability to reason about where to place non-obviously related concepts on an input 3D shape, such as adding clothing to a bare 3D animal model. Our method contextualizes the text description using a neural field and colors the corresponding region of the shape using a probability-weighted blend. Our neural optimization is guided by a pre-trained CLIP encoder, which bypasses the need for any 3D datasets or 3D annotations. Thus, 3D Highlighter is highly flexible, general, and capable of producing localizations on a myriad of input shapes. Our code is publicly available at https://github.com/threedle/3DHighlighter.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 28aa301d-e02a-4139-9baa-5e71c554cd88Cited by top-tier papers11
- SATR: Zero-Shot Semantic Segmentation of 3D ShapesAhmed Abdelreheem, Ivan Skorokhodov, Maks Ovsjanikov, Peter WonkaICCV 2023 · 68 citations
- Style2Fab: Functionality-Aware Segmentation for Fabricating Personalized 3D Models with Generative AIFaraz Faruqi, Ahmed Katary, Tarik Hasic, Amira Abdel-Rahman et al.UIST 2023 · 39 citations
- DAE-Net: Deforming Auto-Encoder for fine-grained shape co-segmentationZhiqin Chen, Qimin Chen, Hang Zhou, Hao ZhangSIGGRAPH 2024 · 9 citations
- PatchAlign3D: Local Feature Alignment for Dense 3D Shape UnderstandingSouhail Hadgi, Bingchen Gong, Ramana Sundararaman, Emery Pierson et al.CVPR 2026 · 5 citations
- Artiverse: A Diverse and Physically Grounded Dataset for Articulated ObjectsDenys Iliash, Jiayi Liu, Egor Fokin, Qirui Wu et al.CVPR 2026 · 4 citations
Builds on14
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Fourier Features Let Networks Learn High Frequency Functions in Low Dimensional DomainsMatthew Tancik, Pratul P. Srinivasan, Ben Mildenhall, Sara Fridovich-Keil et al.NeurIPS 2020 · 4,036 citations
- Language-driven Semantic SegmentationBoyi Li, Kilian Q. Weinberger, Serge J. Belongie, Vladlen Koltun et al.ICLR 2022 · 885 citations
- Decomposing NeRF for Editing via Feature Field DistillationSosuke Kobayashi, Eiichi Matsumoto, Vincent SitzmannNeurIPS 2022 · 479 citations
- DreamFusion: Text-to-3D using 2D DiffusionBen Poole, Ajay Jain, Jonathan T. Barron, Ben MildenhallICLR 2023 · 463 citations
Related papers
- Text2Mesh: Text-Driven Neural Stylization for MeshesOscar Michel, Roi Bar-On, Richard Liu, Sagie Benaim et al.CVPR 2022
- TextDeformer: Geometry Manipulation using Text GuidanceWilliam Gao, Noam Aigerman, Thibault Groueix, Vova Kim et al.SIGGRAPH 2023 · 49 citations
- TANGO: Text-driven Photorealistic and Robust 3D Stylization via Lighting DecompositionYongwei Chen, Rui Chen, Jiabao Lei, Yabin Zhang et al.NeurIPS 2022 · 112 citations
- Zero-Shot Text-Guided Object Generation with Dream FieldsAjay Jain, Ben Mildenhall, Jonathan T. Barron, Pieter Abbeel et al.CVPR 2022 · 361 citations
- CLIP-Forge: Towards Zero-Shot Text-to-Shape GenerationAditya Sanghi, Hang Chu, Joseph G. Lambourne, Ye Wang et al.CVPR 2022 · 206 citations
