SynSin: End-to-End View Synthesis From a Single Image
Olivia Wiles, Georgia Gkioxari, Richard Szeliski, Justin Johnson
摘要
View synthesis allows for the generation of new views of a scene given one or more images. This is challenging; it requires comprehensively understanding the 3D scene from images. As a result, current methods typically use multiple images, train on ground-truth depth, or are limited to synthetic data. We propose a novel end-to-end model for this task using a single image at test time; it is trained on real images without any ground-truth 3D information. To this end, we introduce a novel differentiable point cloud renderer that is used to transform a latent 3D point cloud of features into the target view. The projected features are decoded by our refinement network to inpaint missing regions and generate a realistic output image. The 3D component inside of our generative model allows for interpretable manipulation of the latent feature space at test time, e.g. we can animate trajectories from a single image. Additionally, we can generate high resolution images and generalise to other input resolutions. We outperform baselines and prior work on the Matterport, Replica, and RealEstate10K datasets.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper195
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 被引用 5,687 次
- Plenoxels: Radiance Fields without Neural NetworksSara Fridovich-Keil, Alex Yu, Matthew Tancik, Qinhong Chen 等CVPR 2022 · 被引用 1,237 次
- BARF: Bundle-Adjusting Neural Radiance FieldsChen-Hsuan Lin, Wei-Chiu Ma, Antonio Torralba, Simon LuceyICCV 2021 · 被引用 867 次
- Common Objects in 3D: Large-Scale Learning and Evaluation of Real-life 3D Category ReconstructionJeremy Reizenstein, Roman Shapovalov, Philipp Henzler, Luca Sbordone 等ICCV 2021 · 被引用 686 次
- 2D Gaussian Splatting for Geometrically Accurate Radiance FieldsBinbin Huang, Zehao Yu, Anpei Chen, Andreas Geiger 等SIGGRAPH 2024 · 被引用 660 次
它引用的顶会 Paper7
- Habitat: A Platform for Embodied AI ResearchManolis Savva, Jitendra Malik, Devi Parikh, Dhruv Batra 等ICCV 2019 · 被引用 1,863 次
- Extreme View SynthesisInchang Choi, Orazio Gallo, Alejandro J. Troccoli, Min H. Kim 等ICCV 2019 · 被引用 207 次
- Boundless: Generative Adversarial Networks for Image ExtensionDilip Krishnan, Piotr Teterwak, Aaron Sarna, Aaron Maschinot 等ICCV 2019 · 被引用 129 次
- HoloGAN: Unsupervised Learning of 3D Representations From Natural ImagesThu Nguyen-Phuoc, Chuan Li, Lucas Theis, Christian Richardt 等ICCV 2019 · 被引用 98 次
- Monocular Neural Image Based Rendering With Continuous View ControlJie Song, Xu Chen, Otmar HilligesICCV 2019 · 被引用 85 次
相关 Paper
- Simple and Effective Synthesis of Indoor 3D ScenesJing Yu Koh, Harsh Agrawal, Dhruv Batra, Richard Tucker 等AAAI 2023 · 被引用 40 次
- Worldsheet: Wrapping the World in a 3D Sheet for View Synthesis from a Single ImageRonghang Hu, Nikhila Ravi, Alexander C. Berg, Deepak PathakICCV 2021 · 被引用 97 次
- Ponder: Point Cloud Pre-training via Neural RenderingDi Huang, Sida Peng, Tong He, Honghui Yang 等ICCV 2023 · 被引用 55 次
- MultiDiff: Consistent Novel View Synthesis from a Single ImageNorman Müller, Katja Schwarz, Barbara Rössle, Lorenzo Porzi 等CVPR 2024 · 被引用 14 次
- Uncertainty-Aware Diffusion-Guided Refinement of 3D ScenesSarosij Bose, Arindam Dutta, Sayak Nag, Junge Zhang 等ICCV 2025 · 被引用 3 次
