One-Shot Refiner: Boosting Feed-forward Novel View Synthesis via One-Step Diffusion
Yitong Dong, Qi Zhang, Minchao Jiang, Zhiqiang Wu, Qingnan Fan, Ying Feng, Huaqi Zhang, Hujun Bao, Guofeng Zhang
Abstract
We present a novel framework for high-fidelity novel view synthesis (NVS) from sparse images, addressing key limitations in recent feed-forward 3D Gaussian Splatting (3DGS) methods built on Vision Transformer (ViT) backbones. While ViT-based pipelines offer strong geometric priors, they are often constrained by low-resolution inputs due to computational costs. Moreover, existing generative enhancement methods tend to be 3D-agnostic, resulting in inconsistent structures across views, especially in unseen regions. To overcome these challenges, we design a Dual-Domain Detail Perception Module, which enables handling high-resolution images without being limited by the ViT backbone, and endows Gaussians with additional features to store high-frequency details. We develop a feature-guided diffusion network, which can preserve high-frequency details during the restoration process. We introduce a unified training strategy that enables joint optimization of the ViT-based geometric backbone and the diffusion-based refinement module. Experiments demonstrate that our method can maintain superior generation quality across multiple datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext fcab9270-ef04-4953-b922-8ce9a52f30eeCited by top-tier papers2
- OnlinePG: Online Open-Vocabulary Panoptic Mapping with 3D Gaussian SplattingHongjia Zhai, Qi Zhang, Xiaokun Pan, Xiyu Zhang et al.CVPR 2026 · 3 citations
- TUDSR: Twice Upsampling-Diffusion for Higher Super-ResolutionZhiqiang Wu, Yitong Dong, Xian WeiCVPR 2026
Builds on19
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 5,687 citations
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 2,647 citations
- DUSt3R: Geometric 3D Vision Made EasyShuzhe Wang, Vincent Leroy, Yohann Cabon, Boris Chidlovskii et al.CVPR 2024 · 302 citations
- LangSplat: 3D Language Gaussian SplattingMinghan Qin, Wanhua Li, Jiawei Zhou, Haoqian Wang et al.CVPR 2024 · 164 citations
- ReconFusion: 3D Reconstruction with Diffusion PriorsRundi Wu, Ben Mildenhall, Philipp Henzler, Keunhong Park et al.CVPR 2024 · 137 citations
Related papers
- 3DGS-Enhancer: Enhancing Unbounded 3D Gaussian Splatting with View-consistent 2D Diffusion PriorsXi Liu, Chaoyi Zhou, Siyu HuangNeurIPS 2024 · 127 citations
- Z-Order Transformer for Feed-Forward Gaussian SplattingCan Wang, Lei Liu, Wei Jiang, Dong XuCVPR 2026 · 1 citation
- MuSASplat: Efficient Sparse-View 3D Gaussian Splats via Lightweight Multi-Scale AdaptationMuyu Xu, Fangneng Zhan, Xiaoqin Zhang, Ling Shao et al.AAAI 2026
- GeoQuery: Geometry-Query Diffusion for Sparse-View ReconstructionXiao Cao, Yuze Li, Youmin Zhang, Jiayu Song et al.SIGGRAPH 2026
- PR-IQA: Partial-Reference Image Quality Assessment for Diffusion-Based Novel View SynthesisInseong Choi, Siwoo Lee, Seung-Hun Nam, Soohwan SongCVPR 2026
