MVInpainter: Learning Multi-View Consistent Inpainting to Bridge 2D and 3D Editing
Chenjie Cao, Chaohui Yu, Fan Wang, Xiangyang Xue, Yanwei Fu
Abstract
Novel View Synthesis (NVS) and 3D generation have recently achieved prominent improvements. However, these works mainly focus on confined categories or synthetic 3D assets, which are discouraged from generalizing to challenging in-the-wild scenes and fail to be employed with 2D synthesis directly. Moreover, these methods heavily depended on camera poses, limiting their real-world applications. To overcome these issues, we propose MVInpainter, re-formulating the 3D editing as a multi-view 2D inpainting task. Specifically, MVInpainter partially inpaints multi-view images with the reference guidance rather than intractably generating an entirely novel view from scratch, which largely simplifies the difficulty of in-the-wild NVS and leverages unmasked clues instead of explicit pose conditions. To ensure cross-view consistency, MVInpainter is enhanced by video priors from motion components and appearance guidance from concatenated reference key&value attention. Furthermore, MVInpainter incorporates slot attention to aggregate high-level optical flow features from unmasked regions to control the camera movement with pose-free training and inference. Sufficient scene-level experiments on both object-centric and forward-facing datasets verify the effectiveness of MVInpainter, including diverse tasks, such as multi-view object removal, synthesis, insertion, and replacement. The project page is https://ewrfcas.github.io/MVInpainter/.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 05eb5b3b-14a4-4add-99ec-2e0602d92d19Cited by top-tier papers8
- Omni-3DEdit: Generalized Versatile 3D Editing in One-PassLiyi Chen, Pengfei Wang, Guowen Zhang, Zhiyuan Ma et al.CVPR 2026 · 6 citations
- DiGA3D: Coarse-to-Fine Diffusional Propagation of Geometry and Appearance for Versatile 3D InpaintingJingyi Pan, Dan Xu, Qiong LuoICCV 2025 · 3 citations
- Perspective-Aware 3D Gaussian Inpainting with Multi-View ConsistencyYuxin Cheng, Binxiao Huang, Taiqiang Wu, Wenyong Zhou et al.ICCV 2025 · 1 citation
- Human Interaction-Aware 3D Reconstruction from a Single ImageGwanghyun Kim, Junghun James Kim, Suh Yoon Jeon, Jason Park et al.CVPR 2026
- FreeInsert: Disentangled Text-Guided Object Insertion in 3D Gaussian Scene without Spatial PriorsChenxi Li, Weijie Wang, Qiang Li, Nicu Sebe et al.ACM MM 2025
Builds on70
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu et al.ICLR 2022 · 18,833 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li et al.NeurIPS 2022 · 8,965 citations
Related papers
- LaRP: Efficient Multi-View Inpainting with Latent Reprojection PriorsGaoyang Zhang, Xinguo LiuCVPR 2026
- NViST: In the Wild New View Synthesis from a Single Image with TransformersWonbong Jang, Lourdes AgapitoCVPR 2024
- Free3D: Consistent Novel View Synthesis Without 3D RepresentationChuanxia Zheng, Andrea VedaldiCVPR 2024 · 28 citations
- NVComposer: Boosting Generative Novel View Synthesis with Multiple Sparse and Unposed ImagesLingen Li, Zhaoyang Zhang, Yaowei Li, Jiale Xu et al.CVPR 2025
- SPIn-NeRF: Multiview Segmentation and Perceptual Inpainting with Neural Radiance FieldsAshkan Mirzaei, Tristan Aumentado-Armstrong, Konstantinos G. Derpanis, Jonathan Kelly et al.CVPR 2023
