3D-LATTE: Latent Space 3D Editing from Textual Instructions
Maria Parelli, Michael Oechsle, Michael Niemeyer, Federico Tombari, Andreas Geiger
摘要
Despite the recent success of multi-view diffusion models for text/image-based 3D asset generation, instruction-based editing of 3D assets lacks surprisingly far behind the quality of generation models. The main reason is that recent approaches using 2D priors suffer from view-inconsistent editing signals. Going beyond 2D prior distillation methods and multi-view editing strategies, we propose a training-free editing method that operates within the latent space of a native 3D diffusion model, allowing us to directly manipulate 3D geometry. We guide the edit synthesis by blending 3D attention maps from the generation with the source object. Coupled with geometry-aware regularization guidance, a spectral modulation strategy in the Fourier domain and a refinement step for 3D enhancement, our method outperforms previous 3D editing methods enabling high-fidelity and precise edits across a wide range of shapes and semantic manipulations. Our project webpage is https://mparelli.github.io/3d-latte
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Variation-aware Flexible 3D Gaussian EditingHao Qin, Yukai Sun, Meng Wang, Ming Kong 等ICLR 2026 · 被引用 3 次
- InstructMix2Mix: Consistent Sparse-View Editing Through Multi-View Model PersonalizationDaniel Gilo, Or LitanyCVPR 2026 · 被引用 2 次
它引用的顶会 Paper30
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 被引用 6,759 次
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 被引用 5,687 次
相关 Paper
- Energy-Guided Optimization for Personalized Image Editing with Pretrained Text-to-Image Diffusion ModelsRui Jiang, Xinghe Fu, Guangcong Zheng, Teng Li 等AAAI 2025 · 被引用 2 次
- ShapeUP: Scalable Image-Conditioned 3D EditingInbar Gat, Dana Cohen-Bar, Guy Levy, Elad Richardson 等SIGGRAPH 2026
- Vox-E: Text-guided Voxel Editing of 3D ObjectsEtai Sella, Gal Fiebelman, Peter Hedman, Hadar Averbuch-ElorICCV 2023 · 被引用 122 次
- AnchorFlow: Training-Free 3D Editing via Latent Anchor-Aligned FlowsZhenglin Zhou, Fan Ma, Chengzhuo Gui, Xiaobo Xia 等CVPR 2026 · 被引用 11 次
- Edit360: 2D Image Edits to 3D Assets From Any AngleJunchao Huang, Xinting Hu, Shaoshuai Shi, Zhuotao Tian 等ICCV 2025 · 被引用 2 次
