InterGSEdit: Interactive 3D Gaussian Splatting Editing with 3D Geometry-Consistent Attention Prior
Minghao Wen, Shengjie Wu, Kangkan Wang, Dong Liang
摘要
3D Gaussian Splatting based 3D editing has demonstrated impressive performance in recent years. However, the multi-view editing often exhibits significant local inconsistency, especially in areas of non-rigid deformation, which lead to local artifacts, texture blurring, or semantic variations in edited 3D scenes. We also found that the existing editing methods, which rely entirely on text prompts make the editing process a "one-shot deal", making it difficult for users to control the editing degree flexibly. In response to these challenges, we present InterGSEdit, a novel framework for high-quality 3DGS editing via interactively selecting key views with users' preferences. We propose a CLIP-based Semantic Consistency Selection (CSCS) strategy to adaptively screen a group of semantically consistent reference views for each user-selected key view. Then, the cross-attention maps derived from the reference views are used in a weighted Gaussian Splatting unprojection to construct the 3D Geometry-Consistent Attention Prior (). We project to obtain 3D-constrained attention, which are fused with 2D cross-attention via Attention Fusion Network (AFN). AFN employs an adaptive attention strategy that prioritizes 3D-constrained attention for geometric consistency during early inference, and gradually prioritizes 2D cross-attention maps in diffusion for fine-grained features during the later inference. Extensive experiments demonstrate that InterGSEdit achieves state-of-the-art performance, delivering consistent, high-fidelity 3DGS editing with improved user experience.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- 3D-LATTE: Latent Space 3D Editing from Textual InstructionsMaria Parelli, Michael Oechsle, Michael Niemeyer, Federico Tombari 等CVPR 2026 · 被引用 10 次
- Omni-3DEdit: Generalized Versatile 3D Editing in One-PassLiyi Chen, Pengfei Wang, Guowen Zhang, Zhiyuan Ma 等CVPR 2026 · 被引用 6 次
- SeeThrough3D: Occlusion Aware 3D Control in Text-to-Image GenerationVaibhav Agrawal, Rishubh Parihar, Pradhaan Bhat, Ravi Kiran Sarvadevabhatla 等CVPR 2026 · 被引用 5 次
- ST4R-Splat: Spatio-Temporal Referring Segmentation in 4D Gaussian SplattingYuming Meng, Dong Wu, Hongbin ZhaCVPR 2026
- Catalyst4D: High-Fidelity 3D-to-4D Scene Editing via Dynamic PropagationShifeng Chen, Yihui Li, Jun Liao, Hongyu Yang 等CVPR 2026
它引用的顶会 Paper29
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 被引用 6,759 次
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 被引用 5,687 次
相关 Paper
- ProEdit: Simple Progression is All You Need for High-Quality 3D Scene EditingJun-Kun Chen, Yu-Xiong WangNeurIPS 2024 · 被引用 18 次
- D2Gaussian: Dynamic Control with Discretized 3D View Modeling for Text-Driven 3D Gaussian Splatting EditingYefei Sheng, Jie Wang, Ming Tao, Bing-Kun BaoACM MM 2025 · 被引用 1 次
- Drag Your Gaussian: Effective Drag-Based Editing with Score Distillation for 3D Gaussian SplattingYansong Qu, Dian Chen, Xinyang Li, Xiaofan Li 等SIGGRAPH 2025 · 被引用 14 次
- Personalize Your Gaussian: Consistent 3D Scene Personalization from a Single ImageYuxuan Wang, Xuanyu Yi, Qingshan Xu, Yuan Zhou 等AAAI 2026 · 被引用 2 次
- 3D Gaussian Editing with A Single ImageGuan Luo, Tian-Xing Xu, Ying-Tian Liu, Xiaoxiong Fan 等ACM MM 2024 · 被引用 7 次
