GS-Scale: Unlocking Large-Scale 3D Gaussian Splatting Training via Host Offloading
Donghyun Lee, Dawoon Jeong, Jae W. Lee, Hongil Yoon
Abstract
The advent of 3D Gaussian Splatting has revolutionized graphics rendering by delivering high visual quality and fast rendering speeds. However, training large-scale scenes at high quality remains challenging due to the substantial memory demands required to store parameters, gradients, and optimizer states, which can quickly overwhelm GPU memory. To address these limitations, we propose GS-Scale, a fast and memory-efficient training system for 3D Gaussian Splatting. GS-Scale stores all Gaussians in host memory, transferring only a subset to the GPU on demand for each forward and backward pass. While this dramatically reduces GPU memory usage, it requires frustum culling and optimizer updates to be executed on the CPU, introducing slowdowns due to CPU's limited compute and memory bandwidth. To mitigate this, GS-Scale employs three system-level optimizations: (1) selective offloading of geometric parameters for fast frustum culling, (2) parameter forwarding to pipeline CPU optimizer updates with GPU computation, and (3) deferred optimizer update to minimize unnecessary memory accesses for Gaussians with zero gradients. Our extensive evaluations on large-scale datasets demonstrate that GS-Scale significantly lowers GPU memory demands by 3.3-5.6x, while achieving training speeds comparable to GPU without host offloading. This enables large-scale 3D Gaussian Splatting training on consumer-grade GPUs; for instance, GS-Scale can scale the number of Gaussians from 4 million to 18 million on an RTX 4070 Mobile GPU, leading to 23-35% LPIPS (learned perceptual image patch similarity) improvement.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext dbed39ac-237c-4df2-9ed7-5de7eb551354Cited by top-tier papers2
- Scalable Training of 3D Gaussian Splatting via Out-of-Core OptimizationChonghao Zhong, Shi Linfeng, ChenHua, Tiecheng Sun et al.ICML 2026 · 1 citation
- A LoD of Gaussians: Out-of-Core Training and Rendering for Seamless Ultra-Large Scene ReconstructionFelix Windisch, Thomas Köhler, Lukas Radl, Mattia D'Urso et al.SIGGRAPH 2026
Builds on25
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 4,089 citations
- Plenoxels: Radiance Fields without Neural NetworksSara Fridovich-Keil, Alex Yu, Matthew Tancik, Qinhong Chen et al.CVPR 2022 · 1,237 citations
- FastNeRF: High-Fidelity Neural Rendering at 200FPSStephan J. Garbin, Marek Kowalski, Matthew Johnson, Jamie Shotton et al.ICCV 2021 · 778 citations
- Mega-NeRF: Scalable Construction of Large-Scale NeRFs for Virtual Fly- ThroughsHaithem Turki, Deva Ramanan, Mahadev SatyanarayananCVPR 2022 · 364 citations
- MatrixCity: A Large-scale City Dataset for City-scale Neural Rendering and BeyondYixuan Li, Lihan Jiang, Linning Xu, Yuanbo Xiangli et al.ICCV 2023 · 185 citations
Related papers
- DashGaussian: Optimizing 3D Gaussian Splatting in 200 SecondsYouyu Chen, Junjun Jiang, Kui Jiang, Xiao Tang et al.CVPR 2025
- CLM: Removing the GPU Memory Barrier for 3D Gaussian SplattingHexu Zhao, Xiwen Min, Xiaoteng Liu, Moonjun Gong et al.ASPLOS 2026 · 1 citation
- 3DGS-LM: Faster Gaussian-Splatting Optimization with Levenberg-MarquardtLukas Höllein, Aljaz Bozic, Michael Zollhöfer, Matthias NießnerICCV 2025 · 10 citations
- Local-GS: An Order-Independent Gaussian Splatting Training Accelerator Exploiting Splat LocalityYiyang Sun, Qinzhe Zhi, Yiqi Jing, Le Ye et al.DAC 2025 · 1 citation
- GSArch: Breaking Memory Barriers in 3D Gaussian Splatting Training via Architectural SupportHoushu He, Gang Li, Fangxin Liu, Li Jiang et al.HPCA 2025 · 15 citations
