KeyGS: A Keyframe-Centric Gaussian Splatting Method for Monocular Image Sequences
Keng Wei Chang, Zi-Ming Wang, Shang-Hong Lai
Abstract
Reconstructing high-quality 3D models from sparse 2D images has garnered significant attention in computer vision. Recently, 3D Gaussian Splatting (3DGS) has gained prominence due to its explicit representation with efficient training speed and real-time rendering capabilities. However, existing methods still heavily depend on accurate camera poses for reconstruction. Although some recent approaches attempt to train 3DGS models without the Structure-from-Motion (SfM) preprocessing from monocular video datasets, these methods suffer from prolonged training times, making them impractical for many applications.
In this paper, we present an efficient framework that operates without any depth or matching model. Our approach initially uses SfM to quickly obtain rough camera poses within seconds, and then refines these poses by leveraging the dense representation in 3DGS. This framework effectively addresses the issue of long training times. Additionally, we integrate the densification process with joint refinement and propose a coarse-to-fine frequency-aware densification to reconstruct different levels of details. This approach prevents camera pose estimation from being trapped in local minima or drifting due to high-frequency signals. Our method significantly reduces training time from hours to minutes while achieving more accurate novel view synthesis and camera pose estimation compared to previous methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b25f3186-9893-4656-961c-cd9f21732b85Cited by top-tier papers2
- NopeRoomGS: Indoor 3D Gaussian Splatting Optimization without Camera Pose InputWenbo Li, Yan Xu, Mingde Yao, Fengjie Liang et al.NeurIPS 2025 · 1 citation
- DentalGS: Pose-Free 3D Gaussian Splatting from Five Intraoral Images for Novel View SynthesisHonghao Dai, Yuanfeng Zhou, Guangshun Wei, Zhihao Li et al.AAAI 2026
Builds on22
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 5,687 citations
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 4,089 citations
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 2,647 citations
- PlenOctrees for Real-time Rendering of Neural Radiance FieldsAlex Yu, Ruilong Li, Matthew Tancik, Hao Li et al.ICCV 2021 · 1,284 citations
- BARF: Bundle-Adjusting Neural Radiance FieldsChen-Hsuan Lin, Wei-Chiu Ma, Antonio Torralba, Simon LuceyICCV 2021 · 867 citations
Related papers
- COLMAP-Free 3D Gaussian SplattingYang Fu, Xiaolong Wang, Sifei Liu, Amey Kulkarni et al.CVPR 2024
- Deep Gaussian from Motion: Exploring 3D Geometric Foundation Models for Gaussian SplattingYu Chen, Rolandos Alexandros Potamias, Evangelos Ververas, Jifei Song et al.NeurIPS 2025
- Pose-Free Omnidirectional Gaussian Splatting for 360-Degree Videos with Consistent Depth PriorsChuanqing Zhuang, Xin Lu, Zehui Deng, Zhengda Lu et al.CVPR 2026
- A Construct-Optimize Approach to Sparse View Synthesis without Camera PoseKaiwen Jiang, Yang Fu, Mukund Varma T., Yash Belhe et al.SIGGRAPH 2024 · 20 citations
- No Pose at All: Self-Supervised Pose-Free 3D Gaussian Splatting from Sparse ViewsRanran Huang, Krystian MikolajczykICCV 2025 · 12 citations
