Taming Video Diffusion Prior with Scene-Grounding Guidance for 3D Gaussian Splatting from Sparse Inputs
Yingji Zhong, Zhihao Li, Dave Zhenyu Chen, Lanqing Hong, Dan Xu
Abstract
3DGS with Vanilla Generation 3DGS with Scene-Grounding Generation … (a) Extrapolation (c) Enhancement (b) Occlusion Figure 1. We tackle the critical issues of (a) extrapolation and (b) occlusion in sparse-input 3DGS by leveraging a video diffusion model. Vanilla generation often suffers from inconsistencies within the generated sequences (as highlighted by the yellow arrows), leading to black shadows in the rendered images. In contrast, our scene-grounding generation produces consistent sequences, effectively addressing these issues and enhancing overall quality (c), as indicated by the blue boxes. The numbers refer to PSNR values. Zoom in for better visualization.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers9
- Quantifying and Alleviating Co-Adaptation in Sparse-View 3D Gaussian SplattingKangjie Chen, Yingji Zhong, Zhihao Li, Jiaqi Lin et al.NeurIPS 2025 · 15 citations
- G4Splat: Geometry-Guided Gaussian Splatting with Generative PriorJunfeng Ni, Yixin Chen, Zhifei Yang, Yu Liu et al.ICLR 2026 · 10 citations
- GaussFusion: Improving 3D Reconstruction in the Wild with A Geometry-Informed Video GeneratorLiyuan Zhu, Manjunath Narayana, Michal Stary, Will Hutchcroft et al.CVPR 2026 · 7 citations
- AnchorSplat: Feed-Forward 3D Gaussian Splatting With 3D Geometric PriorsXiaoxue Zhang, Xiaoxu Zheng, Yixuan Yin, Tiao Zhao et al.CVPR 2026 · 6 citations
- Novel View Synthesis from A Few Glimpses via Test-Time Natural Video CompletionYan Xu, Yixing Wang, Stella X. YuNeurIPS 2025 · 4 citations
Builds on40
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 6,759 citations
Related papers
- Hierarchical Masked 3D Diffusion Model for Video OutpaintingFanda Fan, Chaoxu Guo, Litong Gong, Biao Wang et al.ACM MM 2023 · 12 citations
- 3DGS-Enhancer: Enhancing Unbounded 3D Gaussian Splatting with View-consistent 2D Diffusion PriorsXi Liu, Chaoyi Zhou, Siyu HuangNeurIPS 2024 · 127 citations
- DiffVsgg: Diffusion-Driven Online Video Scene Graph GenerationMu Chen, Liulei Li, Wenguan Wang, Yi YangCVPR 2025
- ExPose: Reinforcing Video Generation Models for Extreme Pose EstimationYoungho Yoon, Wonjune Cho, Hyunho Ha, Sujung Kim et al.CVPR 2026
- GenFusion: Closing the Loop between Reconstruction and Generation via VideosSibo Wu, Congrong Xu, Binbin Huang, Andreas Geiger et al.CVPR 2025
