Look at the Sky: Sky-Aware Efficient 3D Gaussian Splatting in the Wild
Yuze Wang, Junyi Wang, Ruicheng Gao, Yansong Qu, Wantong Duan, Shuo Yang, Yue Qi
Abstract
Photos taken in unconstrained tourist environments often present challenges for accurate 3D scene reconstruction due to variable appearances and transient occlusions, which can introduce artifacts in novel view synthesis. Recently, in-the-wild 3D scene reconstruction has been achieved realistic rendering with Neural Radiance Fields (NeRFs). With the advancement of 3D Gaussian Splatting (3DGS), some methods also attempt to reconstruct 3D scenes from unconstrained photo collections and achieve real-time rendering. However, the rapid convergence of 3DGS is misaligned with the slower convergence of neural network-based appearance encoder and transient mask predictor, hindering the reconstruction efficiency. To address this, we propose a novel sky-aware framework for scene reconstruction from unconstrained photo collection using 3DGS. Firstly, we observe that the learnable per-image transient mask predictor in previous work is unnecessary. By introducing a simple yet efficient greedy supervision strategy, we directly utilize the pseudo mask generated by a pretrained semantic segmentation network as the transient mask, thereby achieving more efficient and higher quality in-the-wild 3D scene reconstruction. Secondly, we find that separately estimating appearance embeddings for the sky and building significantly improves reconstruction efficiency and accuracy. We analyze the underlying reasons and introduce a neural sky module to generate diverse skies from latent sky embeddings extract from unconstrained images. Finally, we propose a mutual distillation learning strategy to constrain sky and building appearance embeddings within the same latent space, further enhancing reconstruction efficiency and quality. Extensive experiments on multiple datasets demonstrate that the proposed framework outperforms existing methods in novel view and appearance synthesis, offering superior rendering quality with faster convergence and rendering speed.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get e3a6a4fe-c750-49e8-86f8-70a3a612c2d2Cited by top-tier papers10
- FastVGGT: Fast Visual Geometry TransformerYou Shen, Zhipeng Zhang, Yansong Qu, Xiawu Zheng et al.ICLR 2026 · 73 citations
- PanoVGGT: Feed-Forward 3D Reconstruction from Panoramic ImageryYijing Guo, Mengjun Chao, Luo Wang, Tianyang Zhao et al.CVPR 2026 · 11 citations
- 3DOT: Texture Transfer for 3DGS Objects from a Single Reference ImageXiao Cao, Beibei Lin, Bo Wang, Zhiyong Huang et al.NeurIPS 2025 · 8 citations
- XSpecMesh: Quality-Preserving Auto-Regressive Mesh Generation Acceleration via Multi-Head Speculative DecodingDian Chen, Yansong Qu, Xinyang Li, Ming Li et al.ICML 2026 · 5 citations
- Training-Free Hierarchical Scene Understanding for Gaussian Splatting with Superpoint GraphsShaohui Dai, Yansong Qu, Zheyan Li, Xinyang Li et al.ACM MM 2025 · 3 citations
Related papers
- Wild-GS: Real-Time Novel View Synthesis from Unconstrained Photo CollectionsJiacong Xu, Yiqun Mei, Vishal M. PatelNeurIPS 2024 · 73 citations
- WildGaussians: 3D Gaussian Splatting In the WildJonas Kulhanek, Songyou Peng, Zuzana Kukelova, Marc Pollefeys et al.NeurIPS 2024 · 202 citations
- MS-GS: Multi-Appearance Sparse-View 3D Gaussian Splatting in the WildDeming Li, Kaiwen Jiang, Yutao Tang, Ravi Ramamoorthi et al.NeurIPS 2025 · 7 citations
- 3D Geometry-aware Deformable Gaussian Splatting for Dynamic View SynthesisZhicheng Lu, Xiang Guo, Le Hui, Tianrui Chen et al.CVPR 2024 · 33 citations
- Taking Language Embedded 3D Gaussian Splatting into the WildYuze Wang, Junyi Wang, Yue QiIEEE VR 2026
