InvSculpt: Inverse Sculpting Modeling via Controlled 3D Generation and a Vector Displacement Field
Hengyu Meng, Lanjiong Li, Zhijing Shao, Yingda Yin, Lingting Zhu, Zeyu Hu, Xin Wang, Ligang Liu, Zeyu Wang
Abstract
Inverse sculpting modeling aims to decompose a sculpted mesh into an underlying base shape and reusable geometric details, enabling non-expert users to inherit professional sculpting effort. We present InvSculpt, a novel inverse sculpting framework that decomposes a sculpted mesh into a high-fidelity underlying shape and reusable geometric details represented as a vector displacement field (VDF). Our approach combines semantic priors from text-guided 2D image editing with a 3D rectified flow model to perform inversion-based, mask-free detail removal, recovering an underlying shape that preserves the identity of the source mesh. To represent sculpted details in a lossless and transferable manner, we extract a VDF defined on the surface of the recovered underlying shape and learn a continuous neural representation for geometry-aware transfer. We observe that standard conditional sampling after inversion often suffers from trajectory drift, leading to identity shift and low-frequency distortion. To address this issue, we introduce a trajectory correction strategy that constrains early sampling steps to follow the inversion path, effectively stabilizing subsequent conditional guidance. This design enables robust detail removal and precise extraction of the VDF. Extensive experiments demonstrate that InvSculpt achieves significantly higher-quality mesh decomposition than prior methods and supports a wide range of applications, including geometry redesign and high-fidelity geometric detail transfer.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3827cf18-d56e-4a37-a184-1b79dcc2fbefBuilds on26
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- DreamFusion: Text-to-3D using 2D DiffusionBen Poole, Ajay Jain, Jonathan T. Barron, Ben MildenhallICLR 2023 · 463 citations
- CLIP-NeRF: Text-and-Image Driven Manipulation of Neural Radiance FieldsCan Wang, Menglei Chai, Mingming He, Dongdong Chen et al.CVPR 2022 · 313 citations
- PartField: Learning 3D Feature Fields for Part Segmentation and BeyondMing-Yu Liu, Mikaela Angelina Uy, Donglai Xiang, Hao Su et al.ICCV 2025 · 103 citations
Related papers
- Text2VDM: Text to Vector Displacement Maps for Expressive and Interactive 3D SculptingHengyu Meng, Duotun Wang, Zhijing Shao, Ligang Liu et al.ICCV 2025 · 2 citations
- Vinedresser3D: Towards Agentic Text-guided 3D EditingYankuan Chi, Xiang Li, Zixuan Huang, James M.CVPR 2026
- Geometry-Consistent Neural Shape Representation with Implicit Displacement FieldsYifan Wang, Lukas Rahmann, Olga Sorkine-HornungICLR 2022 · 81 citations
- Delta Rectified Flow Sampling for Text-to-Image EditingGaspard Beaudouin, Minghan Li, Jaeyeon Kim, Sung-Hoon Yoon et al.CVPR 2026 · 4 citations
- InstantEdit: Text-Guided Few-Step Image Editing with Piecewise Rectified FlowYiming Gong, Zhen Zhu, Minjia ZhangICCV 2025
