Look at the Sky: Sky-Aware Efficient 3D Gaussian Splatting in the Wild
Yuze Wang, Junyi Wang, Ruicheng Gao, Yansong Qu, Wantong Duan, Shuo Yang, Yue Qi
摘要
Photos taken in unconstrained tourist environments often present challenges for accurate 3D scene reconstruction due to variable appearances and transient occlusions, which can introduce artifacts in novel view synthesis. Recently, in-the-wild 3D scene reconstruction has been achieved realistic rendering with Neural Radiance Fields (NeRFs). With the advancement of 3D Gaussian Splatting (3DGS), some methods also attempt to reconstruct 3D scenes from unconstrained photo collections and achieve real-time rendering. However, the rapid convergence of 3DGS is misaligned with the slower convergence of neural network-based appearance encoder and transient mask predictor, hindering the reconstruction efficiency. To address this, we propose a novel sky-aware framework for scene reconstruction from unconstrained photo collection using 3DGS. Firstly, we observe that the learnable per-image transient mask predictor in previous work is unnecessary. By introducing a simple yet efficient greedy supervision strategy, we directly utilize the pseudo mask generated by a pretrained semantic segmentation network as the transient mask, thereby achieving more efficient and higher quality in-the-wild 3D scene reconstruction. Secondly, we find that separately estimating appearance embeddings for the sky and building significantly improves reconstruction efficiency and accuracy. We analyze the underlying reasons and introduce a neural sky module to generate diverse skies from latent sky embeddings extract from unconstrained images. Finally, we propose a mutual distillation learning strategy to constrain sky and building appearance embeddings within the same latent space, further enhancing reconstruction efficiency and quality. Extensive experiments on multiple datasets demonstrate that the proposed framework outperforms existing methods in novel view and appearance synthesis, offering superior rendering quality with faster convergence and rendering speed.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper10
- FastVGGT: Fast Visual Geometry TransformerYou Shen, Zhipeng Zhang, Yansong Qu, Xiawu Zheng 等ICLR 2026 · 被引用 73 次
- PanoVGGT: Feed-Forward 3D Reconstruction from Panoramic ImageryYijing Guo, Mengjun Chao, Luo Wang, Tianyang Zhao 等CVPR 2026 · 被引用 11 次
- 3DOT: Texture Transfer for 3DGS Objects from a Single Reference ImageXiao Cao, Beibei Lin, Bo Wang, Zhiyong Huang 等NeurIPS 2025 · 被引用 8 次
- XSpecMesh: Quality-Preserving Auto-Regressive Mesh Generation Acceleration via Multi-Head Speculative DecodingDian Chen, Yansong Qu, Xinyang Li, Ming Li 等ICML 2026 · 被引用 5 次
- Training-Free Hierarchical Scene Understanding for Gaussian Splatting with Superpoint GraphsShaohui Dai, Yansong Qu, Zheyan Li, Xinyang Li 等ACM MM 2025 · 被引用 3 次
相关 Paper
- Wild-GS: Real-Time Novel View Synthesis from Unconstrained Photo CollectionsJiacong Xu, Yiqun Mei, Vishal M. PatelNeurIPS 2024 · 被引用 73 次
- WildGaussians: 3D Gaussian Splatting In the WildJonas Kulhanek, Songyou Peng, Zuzana Kukelova, Marc Pollefeys 等NeurIPS 2024 · 被引用 202 次
- MS-GS: Multi-Appearance Sparse-View 3D Gaussian Splatting in the WildDeming Li, Kaiwen Jiang, Yutao Tang, Ravi Ramamoorthi 等NeurIPS 2025 · 被引用 7 次
- 3D Geometry-aware Deformable Gaussian Splatting for Dynamic View SynthesisZhicheng Lu, Xiang Guo, Le Hui, Tianrui Chen 等CVPR 2024 · 被引用 33 次
- Taking Language Embedded 3D Gaussian Splatting into the WildYuze Wang, Junyi Wang, Yue QiIEEE VR 2026
