Sequence Matters: Harnessing Video Models in 3D Super-Resolution
Hyun-kyu Ko, Dongheok Park, Youngin Park, Byeonghyeon Lee, Juhee Han, Eunbyung Park
Abstract
3D super-resolution aims to reconstruct high-fidelity 3D models from low-resolution (LR) multi-view images. Early studies primarily focused on single-image super-resolution (SISR) models to upsample LR images into high-resolution images. However, these methods often lack view consistency because they operate independently on each image. Although various post-processing techniques have been extensively explored to mitigate these inconsistencies, they have yet to fully resolve the issues. In this paper, we perform a comprehensive study of 3D super-resolution by leveraging video super-resolution (VSR) models. By utilizing VSR models, we ensure a higher degree of spatial consistency and can reference surrounding spatial information, leading to more accurate and detailed reconstructions. Our findings reveal that VSR models can perform remarkably well even on sequences that lack precise spatial alignment. Given this observation, we propose a simple yet practical approach to align LR images without involving fine-tuning or generating `smooth' trajectory from the trained 3D models over LR images. The experimental results show that the surprisingly simple algorithms can achieve the state-of-the-art results of 3D super-resolution tasks on standard benchmark datasets, such as the NeRF-synthetic and Mip-NeRF 360 datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9eb741d2-902b-44e9-9a6a-8e1324a90ac5Cited by top-tier papers3
- SR3R: Rethinking Super-Resolution 3D Reconstruction With Feed-Forward Gaussian SplattingXiang Feng, Xiangbo Wang, Tieshi Zhong, Chengkai Wang et al.CVPR 2026 · 3 citations
- GaussianZoom: Progressive Zoom-in Generative 3D Gaussian Splatting with Geometric and Semantic GuidanceJiale Shi, Jiarui Hu, Zesong Yang, Kaixuan Luan et al.CVPR 2026
- IE-SRGS: An Internal-External Knowledge Fusion Framework for High-Fidelity 3D Gaussian Splatting Super-ResolutionXiang Feng, Tieshi Zhong, Shuo Chang, Weiliu Wang et al.AAAI 2026
Builds on20
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 5,687 citations
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 4,089 citations
- Mip-NeRF: A Multiscale Representation for Anti-Aliasing Neural Radiance FieldsJonathan T. Barron, Ben Mildenhall, Matthew Tancik, Peter Hedman et al.ICCV 2021 · 2,700 citations
- NeuS: Learning Neural Implicit Surfaces by Volume Rendering for Multi-view ReconstructionPeng Wang, Lingjie Liu, Yuan Liu, Christian Theobalt et al.NeurIPS 2021 · 2,500 citations
Related papers
- NeRF-SR: High Quality Neural Radiance Fields using SupersamplingChen Wang, Xian Wu, Yuan-Chen Guo, Song-Hai Zhang et al.ACM MM 2022 · 115 citations
- Bridging Diffusion Models and 3D Representations: A 3D Consistent Super-Resolution FrameworkYi-Ting Chen, Ting-Hsuan Liao, Pengsheng Guo, Alexander Gerhard Schwing et al.ICCV 2025 · 1 citation
- Geometry-Aware Reference Synthesis for Multi-View Image Super-ResolutionRi Cheng, Yuqi Sun, Bo Yan, Weimin Tan et al.ACM MM 2022 · 4 citations
- Cross-Guided Optimization of Radiance Fields with Multi-View Image Super-Resolution for High-Resolution Novel View SynthesisYoungho Yoon, Kuk-Jin YoonCVPR 2023
- DiSR-NeRF: Diffusion-Guided View-Consistent Super-Resolution NeRFJie Long Lee, Chen Li, Gim Hee LeeCVPR 2024
