ViCA-NeRF: View-Consistency-Aware 3D Editing of Neural Radiance Fields
Jiahua Dong, Yu-Xiong Wang
Abstract
We introduce ViCA-NeRF, the first view-consistency-aware method for 3D editing with text instructions. In addition to the implicit neural radiance field (NeRF) modeling, our key insight is to exploit two sources of regularization that explicitly propagate the editing information across different views, thus ensuring multi-view consistency. For geometric regularization, we leverage the depth information derived from NeRF to establish image correspondences between different views. For learned regularization, we align the latent codes in the 2D diffusion model between edited and unedited images, enabling us to edit key views and propagate the update throughout the entire scene. Incorporating these two strategies, our ViCA-NeRF operates in two stages. In the initial stage, we blend edits from different views to create a preliminary 3D edit. This is followed by a second stage of NeRF training, dedicated to further refining the scene's appearance. Experimental results demonstrate that ViCA-NeRF provides more flexible, efficient (3 times faster) editing with higher levels of consistency and details, compared with the state of the art. Our code is available at: https://dongjiahua.github.io/VICA-NeRF .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1b4ddd3d-898e-4f5e-a062-d9ed6170fd67Cited by top-tier papers27
- ProEdit: Simple Progression is All You Need for High-Quality 3D Scene EditingJun-Kun Chen, Yu-Xiong WangNeurIPS 2024 · 18 citations
- CL-Splats: Continual Learning of Gaussian Splatting with Local OptimizationJan Ackermann, Jonas Kulhanek, Shengqu Cai, Haofei Xu et al.ICCV 2025 · 13 citations
- 3D Gaussian Editing with A Single ImageGuan Luo, Tian-Xing Xu, Ying-Tian Liu, Xiaoxiong Fan et al.ACM MM 2024 · 7 citations
- In-N-Out: Lifting 2D Diffusion Prior for 3D Object Removal via Tuning-Free Latents AlignmentDongting Hu, Huan Fu, Jiaxian Guo, Liuhua Peng et al.NeurIPS 2024 · 6 citations
- Instruct 4D-to-4D: Editing 4D Scenes as Pseudo-3D Scenes Using 2D DiffusionLinzhan Mou, Jun-Kun Chen, Yu-Xiong WangCVPR 2024 · 6 citations
Builds on26
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 4,089 citations
- SDEdit: Guided Image Synthesis and Editing with Stochastic Differential EquationsChenlin Meng, Yutong He, Yang Song, Jiaming Song et al.ICLR 2022 · 2,128 citations
Related papers
- Instruct-NeRF2NeRF: Editing 3D Scenes with InstructionsAyaan Haque, Matthew Tancik, Alexei A. Efros, Aleksander Holynski et al.ICCV 2023 · 544 citations
- InstructPix2NeRF: Instructed 3D Portrait Editing from a Single ImageJianhui Li, Shilong Liu, Zidong Liu, Yikai Wang et al.ICLR 2024 · 12 citations
- SIGNeRF: Scene Integrated Generation for Neural Radiance FieldsJan-Niklas Dihlmann, Andreas Engelhardt, Hendrik P. A. LenschCVPR 2024
- Language-driven Object Fusion into Neural Radiance Fields with Pose-Conditioned Dataset UpdatesKa-Chun Shum, Jaeyeon Kim, Binh-Son Hua, Duc Thanh Nguyen et al.CVPR 2024 · 7 citations
- SINE: Semantic-driven Image-based NeRF Editing with Prior-guided Editing FieldChong Bao, Yinda Zhang, Bangbang Yang, Tianxing Fan et al.CVPR 2023
