No Pose at All: Self-Supervised Pose-Free 3D Gaussian Splatting from Sparse Views
Ranran Huang, Krystian Mikolajczyk
摘要
We introduce SPFSplat, an efficient framework for 3D Gaussian splatting from sparse multi-view images, requiring no ground-truth poses during training or inference. It employs a shared feature extraction backbone, enabling simultaneous prediction of 3D Gaussian primitives and camera poses in a canonical space from unposed inputs within a single feed-forward step. Alongside the rendering loss based on estimated novel-view poses, a reprojection loss is integrated to enforce the learning of pixel-aligned Gaussian primitives for enhanced geometric constraints. This pose-free training paradigm and efficient one-step feed-forward design make SPFSplat well-suited for practical applications. Remarkably, despite the absence of pose supervision, SPFSplat achieves state-of-the-art performance in novel view synthesis even under significant viewpoint changes and limited image overlap. It also surpasses recent methods trained with geometry priors in relative pose estimation. Code and trained models are available on our project page: https://ranrhuang.github.io/spfsplat/.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- E-RayZer: Self-supervised 3D Reconstruction as Spatial Visual Pre-trainingQitao Zhao, Hao Tan, Qianqian Wang, Sai Bi 等CVPR 2026 · 被引用 24 次
- True Self-Supervised Novel View Synthesis is TransferableThomas W. Mitchel, Hyunwoo Ryu, Vincent SitzmannICLR 2026 · 被引用 13 次
- TokenSplat: Token-aligned 3D Gaussian Splatting for Feed-forward Pose-free ReconstructionYihui Li, Chengxin Lv, Zichen Tang, Hongyu Yang 等CVPR 2026 · 被引用 13 次
- The Less You Depend, The More You Learn: Synthesizing Novel Views from Sparse, Unposed Images without Any 3D KnowledgeHaoru Wang, Kai Ye, Minghan Qin, Yangyan Li 等ICLR 2026 · 被引用 11 次
- Off The Grid: Detection of Primitives for Feed-Forward 3D Gaussian SplattingArthur Moreau, Richard Shaw, Michal Nazarczuk, Jisu Shin 等CVPR 2026 · 被引用 10 次
它引用的顶会 Paper27
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 被引用 5,687 次
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 被引用 2,647 次
- MVSNeRF: Fast Generalizable Radiance Field Reconstruction from Multi-View StereoAnpei Chen, Zexiang Xu, Fuqiang Zhao, Xiaoshuai Zhang 等ICCV 2021 · 被引用 1,024 次
- LightGlue: Local Feature Matching at Light SpeedPhilipp Lindenberger, Paul-Edouard Sarlin, Marc PollefeysICCV 2023 · 被引用 936 次
相关 Paper
- No Pose, No Problem: Surprisingly Simple 3D Gaussian Splats from Sparse Unposed ImagesBotao Ye, Sifei Liu, Haofei Xu, Xueting Li 等ICLR 2025
- LongSplat: Robust Unposed 3D Gaussian Splatting for Casual Long VideosChin-Yang Lin, Cheng Sun, Fu-En Yang, Min-Hung Chen 等ICCV 2025 · 被引用 4 次
- A Construct-Optimize Approach to Sparse View Synthesis without Camera PoseKaiwen Jiang, Yang Fu, Mukund Varma T., Yash Belhe 等SIGGRAPH 2024 · 被引用 20 次
- COLMAP-Free 3D Gaussian SplattingYang Fu, Xiaolong Wang, Sifei Liu, Amey Kulkarni 等CVPR 2024
- FreeSplat: Generalizable 3D Gaussian Splatting Towards Free View Synthesis of Indoor ScenesYunsong Wang, Tianxin Huang, Hanlin Chen, Gim Hee LeeNeurIPS 2024 · 被引用 112 次
