Enhanced Stable View Synthesis
Nishant Jain, Suryansh Kumar, Luc Van Gool
摘要
We introduce an approach to enhance the novel view synthesis from images taken from a freely moving camera. The introduced approach focuses on outdoor scenes where recovering accurate geometric scaffold and camera pose is challenging, leading to inferior results using the state-ofthe-art stable view synthesis (SVS) method. SVS and related methods fail for outdoor scenes primarily due to (i) overrelying on the multiview stereo (MVS) for geometric scaffold recovery and (ii) assuming COLMAP computed camera poses as the best possible estimates, despite it being wellstudied that MVS 3D reconstruction accuracy is limited to scene disparity and camera-pose accuracy is sensitive to key-point correspondence selection. This work proposes a principled way to enhance novel view synthesis solutions drawing inspiration from the basics of multiple view geometry. By leveraging the complementary behavior of MVS and monocular depth, we arrive at a better scene depth per view for nearby and far points, respectively. Moreover, our approach jointly refines camera poses with image-based rendering via multiple rotation averaging graph optimization. The recovered scene depth and the camera-pose help better view-dependent on-surface feature aggregation of the entire scene. Extensive evaluation of our approach on the popular benchmark dataset, such as Tanks and Temples, shows substantial improvement in view synthesis results compared to the prior art. For instance, our method shows 1.5 dB of PSNR improvement on the Tank and Temples. Similar statistics are observed when tested on other benchmark datasets such as FVS, Mip-NeRF 360, and DTU.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- A Hierarchical 3D Gaussian Representation for Real-Time Rendering of Very Large DatasetsBernhard Kerbl, Andreas Meuleman, Georgios Kopanas, Michael Wimmer 等SIGGRAPH 2024 · 被引用 180 次
- Stereo Risk: A Continuous Modeling Approach to Stereo MatchingCe Liu, Suryansh Kumar, Shuhang Gu, Radu Timofte 等ICML 2024 · 被引用 8 次
- InsertNeRF: Instilling Generalizability into NeRF with HyperNet ModulesYanqi Bao, Tianyu Ding, Jing Huo, Wenbin Li 等ICLR 2024 · 被引用 7 次
它引用的顶会 Paper10
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 被引用 2,647 次
- Mip-NeRF 360: Unbounded Anti-Aliased Neural Radiance FieldsJonathan T. Barron, Ben Mildenhall, Dor Verbin, Pratul P. Srinivasan 等CVPR 2022 · 被引用 1,603 次
- Common Objects in 3D: Large-Scale Learning and Evaluation of Real-life 3D Category ReconstructionJeremy Reizenstein, Roman Shapovalov, Philipp Henzler, Luca Sbordone 等ICCV 2021 · 被引用 686 次
- Point-NeRF: Point-based Neural Radiance FieldsQiangeng Xu, Zexiang Xu, Julien Philip, Sai Bi 等CVPR 2022 · 被引用 510 次
- Omnidata: A Scalable Pipeline for Making Multi-Task Mid-Level Vision Datasets from 3D ScansAinaz Eftekhar, Alexander Sax, Jitendra Malik, Amir ZamirICCV 2021 · 被引用 422 次
相关 Paper
- Stable View SynthesisGernot Riegler, Vladlen KoltunCVPR 2021
- Urban Radiance FieldsKonstantinos Rematas, Andrew Liu, Pratul P. Srinivasan, Jonathan T. Barron 等CVPR 2022
- Generalizable Novel-View Synthesis Using a Stereo CameraHaechan Lee, Wonjoon Jin, Seung-Hwan Baek, Sunghyun ChoCVPR 2024
- ZeroNVS: Zero-Shot 360-Degree View Synthesis from a Single ImageKyle Sargent, Zizhang Li, Tanmay Shah, Charles Herrmann 等CVPR 2024 · 被引用 45 次
- MuGS: Multi-Baseline Generalizable Gaussian Splatting ReconstructionYaopeng Lou, Li Shen, Tianqi Liu, Jiaqi Li 等ICCV 2025 · 被引用 1 次
