3D Mesh Editing Using Masked LRMs
Will Gao, Dilin Wang, Yuchen Fan, Aljaz Bozic, Tuur Stuyck, Zhengqin Li, Zhao Dong, Rakesh Ranjan, Nikolaos Sarafianos
Abstract
We present a novel approach to shape editing, building on recent progress in 3D reconstruction from multi-view images. We formulate shape editing as a conditional reconstruction problem, where the model must reconstruct the input shape with the exception of a specified 3D region, in which the geometry should be generated from the conditional signal. To this end, we train a conditional Large Reconstruction Model (LRM) for masked reconstruction, using multi-view consistent masks rendered from a randomly generated 3D occlusion, and using one clean viewpoint as the conditional signal. During inference, we manually define a 3D region to edit and provide an edited image from a canonical viewpoint to fill that region. We demonstrate that, in just a single forward pass, our method not only preserves the input geometry in the unmasked region through reconstruction capabilities on par with SoTA, but is also expressive enough to perform a variety of mesh edits from a single image guidance that past works struggle with, while being 2 -10× faster than the top-performing prior work.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 88b3db06-a1b7-44b7-aabe-723729b09c45Cited by top-tier papers5
- CMD: Controllable Multiview Diffusion for 3D Editing and Progressive GenerationPeng Li, Suizhi Ma, Jialiang Chen, Yuan Liu et al.SIGGRAPH 2025 · 8 citations
- Easy3E: Feed-Forward 3D Asset Editing via Rectified Voxel FlowShimin Hu, Yuanyi Wei, Fei Zha, Yudong Guo et al.CVPR 2026 · 7 citations
- MeshPad: Interactive Sketch-Conditioned Artist-Reminiscent Mesh Generation and EditingHaoxuan Li, Ziya Erkoç, Lei Li, Daniele Sirigatti et al.ICCV 2025 · 5 citations
- CraftMesh: High-Fidelity Generative Mesh Manipulation via Poisson Seamless FusionJames Jincheng Hu, Yuxiao Wu, Youcheng Cai, Ligang LiuCVPR 2026 · 3 citations
- Vinedresser3D: Towards Agentic Text-guided 3D EditingYankuan Chi, Xiang Li, Zixuan Huang, James M.CVPR 2026
Builds on41
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- BEiT: BERT Pre-Training of Image TransformersHangbo Bao, Li Dong, Songhao Piao, Furu WeiICLR 2022 · 3,632 citations
- Generative Pretraining From PixelsMark Chen, Alec Radford, Rewon Child, Jeffrey Wu et al.ICML 2020 · 1,773 citations
- Zero-1-to-3: Zero-shot One Image to 3D ObjectRuoshi Liu, Rundi Wu, Basile Van Hoorick, Pavel Tokmakov et al.ICCV 2023 · 1,662 citations
- Volume Rendering of Neural Implicit SurfacesLior Yariv, Jiatao Gu, Yoni Kasten, Yaron LipmanNeurIPS 2021 · 1,421 citations
Related papers
- VecSet-Edit: Unleashing Pre-trained LRM for Mesh Editing from Single ImageTeng-Fang Hsiao, Bo-Kai Ruan, Yu-Lun Liu, Hong-Han ShuaiSIGGRAPH 2026 · 1 citation
- Amodal3R: Amodal 3D Reconstruction from Occluded 2D ImagesTianhao Wu, Chuanxia Zheng, Frank Guan, Andrea Vedaldi et al.ICCV 2025 · 9 citations
- Blended Point Cloud Diffusion for Localized Text-Guided Shape EditingEtai Sella, Noam Atia, Ron Mokady, Hadar Averbuch-ElorICCV 2025 · 4 citations
- Instant3dit: Multiview Inpainting for Fast Editing of 3D ObjectsAmir Barda, Matheus Gadelha, Vladimir G. Kim, Noam Aigerman et al.CVPR 2025
- Nano3D: A Training-Free Approach for Efficient 3D Editing Without MasksJunliang Ye, Shenghao Xie, Ruowen Zhao, Zhengyi Wang et al.ICLR 2026 · 32 citations
