Z-Order Transformer for Feed-Forward Gaussian Splatting
Can Wang, Lei Liu, Wei Jiang, Dong Xu
Abstract
Recent advances in 3D Gaussian Splatting (3DGS) have enabled significant progress in photorealistic novel view synthesis. However, traditional 3DGS relies on a slow, iterative optimization process, which limits its use in scenarios demanding real-time results. To overcome this bottleneck, recent feed-forward methods aim to predict Gaussian attributes directly from images, but they often struggle with the redundancy of Gaussian primitives and rendering quality. In this paper, we introduce a transformer-based architecture specifically designed for feed-forward Gaussian Splatting. Our key insight is that spatial and semantic relationships among Gaussians can be effectively captured through a sparse attention mechanism, enabled by a Z-order strategy that organizes the unstructured Gaussian set into a spatially coherent sequence. Furthermore, we incorporate this Z-order strategy to adaptively suppress redundancy while preserving critical structural details. This allows the transformer to efficiently model context, compress Gaussian primitives, and predict Gaussian attributes in a single forward pass. Comprehensive experiments demonstrate that our method achieves fast and high-quality novel view synthesis with fewer Gaussian primitives.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ff18577e-ac45-48b9-94ab-b61cd3ccf63cBuilds on29
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 5,687 citations
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 4,089 citations
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 2,647 citations
- Depth Anything V2Lihe Yang, Bingyi Kang, Zilong Huang, Zhen Zhao et al.NeurIPS 2024 · 2,305 citations
- FlashAttention-3: Fast and Accurate Attention with Asynchrony and Low-precisionJay Shah, Ganesh Bikshandi, Ying Zhang, Vijay Thakkar et al.NeurIPS 2024 · 727 citations
Related papers
- TokenGS: Decoupling 3D Gaussian Prediction from Pixels with Learnable TokensJiawei Ren, Michal J. Tyszkiewicz, Jiahui Huang, Zan GojcicCVPR 2026 · 13 citations
- Off The Grid: Detection of Primitives for Feed-Forward 3D Gaussian SplattingArthur Moreau, Richard Shaw, Michal Nazarczuk, Jisu Shin et al.CVPR 2026 · 10 citations
- Fast Feedforward 3D Gaussian Splatting CompressionYihang Chen, Qianyi Wu, Mengyao Li, Weiyao Lin et al.ICLR 2025 · 1 citation
- LeanGaussian: Breaking Pixel or Point Cloud Correspondence in Modeling 3D GaussiansJiamin Wu, Kenkun Liu, Han Gao, Xiaoke Jiang et al.CVPR 2025
- ZPressor: Bottleneck-Aware Compression for Scalable Feed-Forward 3DGSWeijie Wang, Donny Y. Chen, Zeyu Zhang, Duochao Shi et al.NeurIPS 2025 · 31 citations
