Instruct-NeRF2NeRF: Editing 3D Scenes with Instructions
Ayaan Haque, Matthew Tancik, Alexei A. Efros, Aleksander Holynski, Angjoo Kanazawa
Abstract
We propose a method for editing NeRF scenes with text-instructions. Given a NeRF of a scene and the collection of images used to reconstruct it, our method uses an image-conditioned diffusion model (InstructPix2Pix) to iteratively edit the input images while optimizing the underlying scene, resulting in an optimized 3D scene that respects the edit instruction. We demonstrate that our proposed method is able to edit large-scale, real-world scenes, and is able to accomplish more realistic, targeted edits than prior work. Result videos can be found on the project website: https://instruct-nerf2nerf.github.io.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d0ba1b5e-7883-4e7c-b237-1de65bc6d123Cited by top-tier papers186
- CAT3D: Create Anything in 3D with Multi-View Diffusion ModelsRuiqi Gao, Aleksander Holynski, Philipp Henzler, Arthur Brussee et al.NeurIPS 2024 · 490 citations
- VR-GS: A Physical Dynamics-Aware Interactive Gaussian Splatting System in Virtual RealityYing Jiang, Chang Yu, Tianyi Xie, Xuan Li et al.SIGGRAPH 2024 · 153 citations
- Text-to-3D with Classifier Score DistillationXin Yu, Yuan-Chen Guo, Yangguang Li, Ding Liang et al.ICLR 2024 · 132 citations
- LLMR: Real-time Prompting of Interactive Worlds using Large Language ModelsFernanda De La Torre, Cathy Mengying Fang, Han Huang, Andrzej Banburski-Fahey et al.CHI 2024 · 124 citations
- HIFA: High-fidelity Text-to-3D Generation with Advanced Diffusion GuidanceJunzhe Zhu, Peiye Zhuang, Sanmi KoyejoICLR 2024 · 113 citations
Builds on32
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida et al.NeurIPS 2022 · 24,707 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li et al.NeurIPS 2022 · 8,965 citations
Related papers
- InstructPix2NeRF: Instructed 3D Portrait Editing from a Single ImageJianhui Li, Shilong Liu, Zidong Liu, Yikai Wang et al.ICLR 2024 · 12 citations
- ViCA-NeRF: View-Consistency-Aware 3D Editing of Neural Radiance FieldsJiahua Dong, Yu-Xiong WangNeurIPS 2023 · 97 citations
- Language-driven Object Fusion into Neural Radiance Fields with Pose-Conditioned Dataset UpdatesKa-Chun Shum, Jaeyeon Kim, Binh-Son Hua, Duc Thanh Nguyen et al.CVPR 2024 · 7 citations
- GaussianEditor: Editing 3D Gaussians Delicately with Text InstructionsJunjie Wang, Jiemin Fang, Xiaopeng Zhang, Lingxi Xie et al.CVPR 2024 · 65 citations
- SKED: Sketch-guided Text-based 3D EditingAryan Mikaeili, Or Perel, Mehdi Safaee, Daniel Cohen-Or et al.ICCV 2023 · 83 citations
