Efficient Dynamic Scene Editing via 4D Gaussian-based Static-Dynamic Separation
JooHyun Kwon, Hanbyel Cho, Junmo Kim
摘要
Recent 4D dynamic scene editing methods require editing thousands of 2D images used for dynamic scene synthesis and updating the entire scene with additional training loops, resulting in several hours of processing to edit a single dynamic scene. Therefore, these methods are not scalable with respect to the temporal dimension of the dynamic scene (i.e., the number of timesteps). In this work, we propose Instruct-4DGS, an efficient dynamic scene editing method that is more scalable in terms of temporal dimension. To achieve computational efficiency, we leverage a 4D Gaussian representation that models a 4D dynamic scene by combining static 3D Gaussians with a Hexplane-based deformation field, which captures dynamic information. We then perform editing solely on the static 3D Gaussians, which is the minimal but sufficient component required for visual editing. To resolve the misalignment between the edited 3D Gaussians and the deformation field, which may arise from the editing process, we introduce a refinement stage using a score distillation mechanism. Extensive editing results demonstrate that Instruct-4DGS is efficient, reducing editing time by more than half compared to existing methods while achieving high-quality edits that better follow user instructions.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Let it Snow! Animating 3D Gaussian Scenes with Dynamic Weather Effects via Physics-Guided Score DistillationGal Fiebelman, Hadar Averbuch-Elor, Sagie BenaimCVPR 2026 · 被引用 6 次
- Dynamic-eDiTor: Training-Free Text-Driven 4D Scene Editing with Multimodal Diffusion TransformerDong In Lee, Hyungjun Doh, Seunggeun Chi, Runlin Duan 等CVPR 2026 · 被引用 3 次
- Catalyst4D: High-Fidelity 3D-to-4D Scene Editing via Dynamic PropagationShifeng Chen, Yihui Li, Jun Liao, Hongyu Yang 等CVPR 2026
- LidarPainter: One-Step Away from Any Lidar View to Novel GuidanceYuzhou Ji, Ke Ma, Hong Cai, Anchun Zhang 等AAAI 2026
- VideoWeaver: Multimodal Multi-View Video-to-Video Transfer for Embodied AgentsGeorge Eskandar, Fengyi Shen, Mohammad Altillawi, Dong Chen 等CVPR 2026
它引用的顶会 Paper46
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 被引用 6,759 次
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 被引用 5,687 次
相关 Paper
- 4D Gaussian Splatting for Real-Time Dynamic Scene RenderingGuanjun Wu, Taoran Yi, Jiemin Fang, Lingxi Xie 等CVPR 2024 · 被引用 513 次
- Control4D: Efficient 4D Portrait Editing With TextRuizhi Shao, Jingxiang Sun, Cheng Peng, Zerong Zheng 等CVPR 2024 · 被引用 17 次
- ST-4DGS: Spatial-Temporally Consistent 4D Gaussian Splatting for Efficient Dynamic Scene RenderingDeqi Li, Shi-Sheng Huang, Zhiyuan Lu, Xinran Duan 等SIGGRAPH 2024 · 被引用 33 次
- Align Your Gaussians: Text-to-4D with Dynamic 3D Gaussians and Composed Diffusion ModelsHuan Ling, Seung Wook Kim, Antonio Torralba, Sanja Fidler 等CVPR 2024
- Grid4D: 4D Decomposed Hash Encoding for High-Fidelity Dynamic Gaussian SplattingJiawei Xu, Zexin Fan, Jian Yang, Jin XieNeurIPS 2024 · 被引用 64 次
