SplatFormer: Point Transformer for Robust 3D Gaussian Splatting
Yutong Chen, Marko Mihajlovic, Xiyi Chen, Yiming Wang, Sergey Prokudin, Siyu Tang
摘要
3D Gaussian Splatting (3DGS) has recently transformed photorealistic reconstruction, achieving high visual fidelity and real-time performance. However, rendering quality significantly deteriorates when test views deviate from the camera angles used during training, posing a major challenge for applications in immersive free-viewpoint rendering and navigation. In this work, we conduct a comprehensive evaluation of 3DGS and related novel view synthesis methods under out-of-distribution (OOD) test camera scenarios. By creating diverse test cases with synthetic and real-world datasets, we demonstrate that most existing methods, including those incorporating various regularization techniques and data-driven priors, struggle to generalize effectively to OOD views. To address this limitation, we introduce SplatFormer, the first point transformer model specifically designed to operate on Gaussian splats. SplatFormer takes as input an initial 3DGS set optimized under limited training views and refines it in a single forward pass, effectively removing potential artifacts in OOD test views. To our knowledge, this is the first successful application of point transformers directly on 3DGS sets, surpassing the limitations of previous multi-scene training methods, which could handle only a restricted number of input views during inference. Our model significantly improves rendering quality under extreme novel views, achieving state-of-the-art performance in these challenging scenarios and outperforming various 3DGS regularization techniques, multi-scene models tailored for sparse view synthesis, and diffusion-based frameworks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- BevSplat: Resolving Height Ambiguity via Feature-Based Gaussian Primitives for Weakly-Supervised Cross-View LocalizationQiwei Wang, Shaoxun Wu, Yujiao ShiNeurIPS 2025 · 被引用 10 次
- SparseSplat: Towards Applicable Feed-Forward 3D Gaussian Splatting with Pixel-Unaligned PredictionZicheng Zhang, Xiangting Meng, Ke Wu, Wenchao DingCVPR 2026 · 被引用 7 次
- GaussFusion: Improving 3D Reconstruction in the Wild with A Geometry-Informed Video GeneratorLiyuan Zhu, Manjunath Narayana, Michal Stary, Will Hutchcroft 等CVPR 2026 · 被引用 7 次
- How Many Tokens Do 3D Point Cloud Transformer Architectures Really Need?Tuan Anh Tran, Duy M. H. Nguyen, Hoai-Chau Tran, Michael Barz 等NeurIPS 2025 · 被引用 5 次
- UFO: Unifying Feed-Forward and Optimization-based Methods for Large Driving Scene ModelingKaiyuan Tan, Yingying Shen, Ziyue Zhu, Mingfei Tu 等CVPR 2026 · 被引用 4 次
它引用的顶会 Paper38
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 被引用 5,687 次
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 被引用 4,089 次
- Zero-1-to-3: Zero-shot One Image to 3D ObjectRuoshi Liu, Rundi Wu, Basile Van Hoorick, Pavel Tokmakov 等ICCV 2023 · 被引用 1,662 次
相关 Paper
- Z-Order Transformer for Feed-Forward Gaussian SplattingCan Wang, Lei Liu, Wei Jiang, Dong XuCVPR 2026 · 被引用 1 次
- DcSplat: Dual-Constraint Human Gaussian Splatting with Latent Multi-View ConsistencyTengfei Xiao, Yue Wu, Zhigang Gao, Yongzhe Yuan 等AAAI 2026
- COSMOS: Coherent SuperGaussian Modeling with Spatial Priors for Sparse-View 3D SplattingChaeyoung Jeong, Kwangsu KimAAAI 2026
- Perceptual Quality Assessment of 3D Gaussian Splatting: A Subjective Dataset and Prediction MetricZhaolin Wan, Yining Diao, Jingqi Xu, Hao Wang 等AAAI 2026 · 被引用 1 次
- FreeSplat: Generalizable 3D Gaussian Splatting Towards Free View Synthesis of Indoor ScenesYunsong Wang, Tianxin Huang, Hanlin Chen, Gim Hee LeeNeurIPS 2024 · 被引用 112 次
