3D Mesh Editing Using Masked LRMs
Will Gao, Dilin Wang, Yuchen Fan, Aljaz Bozic, Tuur Stuyck, Zhengqin Li, Zhao Dong, Rakesh Ranjan, Nikolaos Sarafianos
摘要
We present a novel approach to shape editing, building on recent progress in 3D reconstruction from multi-view images. We formulate shape editing as a conditional reconstruction problem, where the model must reconstruct the input shape with the exception of a specified 3D region, in which the geometry should be generated from the conditional signal. To this end, we train a conditional Large Reconstruction Model (LRM) for masked reconstruction, using multi-view consistent masks rendered from a randomly generated 3D occlusion, and using one clean viewpoint as the conditional signal. During inference, we manually define a 3D region to edit and provide an edited image from a canonical viewpoint to fill that region. We demonstrate that, in just a single forward pass, our method not only preserves the input geometry in the unmasked region through reconstruction capabilities on par with SoTA, but is also expressive enough to perform a variety of mesh edits from a single image guidance that past works struggle with, while being 2 -10× faster than the top-performing prior work.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- CMD: Controllable Multiview Diffusion for 3D Editing and Progressive GenerationPeng Li, Suizhi Ma, Jialiang Chen, Yuan Liu 等SIGGRAPH 2025 · 被引用 8 次
- Easy3E: Feed-Forward 3D Asset Editing via Rectified Voxel FlowShimin Hu, Yuanyi Wei, Fei Zha, Yudong Guo 等CVPR 2026 · 被引用 7 次
- MeshPad: Interactive Sketch-Conditioned Artist-Reminiscent Mesh Generation and EditingHaoxuan Li, Ziya Erkoç, Lei Li, Daniele Sirigatti 等ICCV 2025 · 被引用 5 次
- CraftMesh: High-Fidelity Generative Mesh Manipulation via Poisson Seamless FusionJames Jincheng Hu, Yuxiao Wu, Youcheng Cai, Ligang LiuCVPR 2026 · 被引用 3 次
- Vinedresser3D: Towards Agentic Text-guided 3D EditingYankuan Chi, Xiang Li, Zixuan Huang, James M.CVPR 2026
它引用的顶会 Paper41
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- BEiT: BERT Pre-Training of Image TransformersHangbo Bao, Li Dong, Songhao Piao, Furu WeiICLR 2022 · 被引用 3,632 次
- Generative Pretraining From PixelsMark Chen, Alec Radford, Rewon Child, Jeffrey Wu 等ICML 2020 · 被引用 1,773 次
- Zero-1-to-3: Zero-shot One Image to 3D ObjectRuoshi Liu, Rundi Wu, Basile Van Hoorick, Pavel Tokmakov 等ICCV 2023 · 被引用 1,662 次
- Volume Rendering of Neural Implicit SurfacesLior Yariv, Jiatao Gu, Yoni Kasten, Yaron LipmanNeurIPS 2021 · 被引用 1,421 次
相关 Paper
- VecSet-Edit: Unleashing Pre-trained LRM for Mesh Editing from Single ImageTeng-Fang Hsiao, Bo-Kai Ruan, Yu-Lun Liu, Hong-Han ShuaiSIGGRAPH 2026 · 被引用 1 次
- Amodal3R: Amodal 3D Reconstruction from Occluded 2D ImagesTianhao Wu, Chuanxia Zheng, Frank Guan, Andrea Vedaldi 等ICCV 2025 · 被引用 9 次
- Blended Point Cloud Diffusion for Localized Text-Guided Shape EditingEtai Sella, Noam Atia, Ron Mokady, Hadar Averbuch-ElorICCV 2025 · 被引用 4 次
- Instant3dit: Multiview Inpainting for Fast Editing of 3D ObjectsAmir Barda, Matheus Gadelha, Vladimir G. Kim, Noam Aigerman 等CVPR 2025
- Nano3D: A Training-Free Approach for Efficient 3D Editing Without MasksJunliang Ye, Shenghao Xie, Ruowen Zhao, Zhengyi Wang 等ICLR 2026 · 被引用 32 次
