SynSin: End-to-End View Synthesis From a Single Image
Olivia Wiles, Georgia Gkioxari, Richard Szeliski, Justin Johnson
Abstract
View synthesis allows for the generation of new views of a scene given one or more images. This is challenging; it requires comprehensively understanding the 3D scene from images. As a result, current methods typically use multiple images, train on ground-truth depth, or are limited to synthetic data. We propose a novel end-to-end model for this task using a single image at test time; it is trained on real images without any ground-truth 3D information. To this end, we introduce a novel differentiable point cloud renderer that is used to transform a latent 3D point cloud of features into the target view. The projected features are decoded by our refinement network to inpaint missing regions and generate a realistic output image. The 3D component inside of our generative model allows for interpretable manipulation of the latent feature space at test time, e.g. we can animate trajectories from a single image. Additionally, we can generate high resolution images and generalise to other input resolutions. We outperform baselines and prior work on the Matterport, Replica, and RealEstate10K datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e23b61cf-09dd-45a0-b16c-91c1d411eb03Cited by top-tier papers195
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 5,687 citations
- Plenoxels: Radiance Fields without Neural NetworksSara Fridovich-Keil, Alex Yu, Matthew Tancik, Qinhong Chen et al.CVPR 2022 · 1,237 citations
- BARF: Bundle-Adjusting Neural Radiance FieldsChen-Hsuan Lin, Wei-Chiu Ma, Antonio Torralba, Simon LuceyICCV 2021 · 867 citations
- Common Objects in 3D: Large-Scale Learning and Evaluation of Real-life 3D Category ReconstructionJeremy Reizenstein, Roman Shapovalov, Philipp Henzler, Luca Sbordone et al.ICCV 2021 · 686 citations
- 2D Gaussian Splatting for Geometrically Accurate Radiance FieldsBinbin Huang, Zehao Yu, Anpei Chen, Andreas Geiger et al.SIGGRAPH 2024 · 660 citations
Builds on7
- Habitat: A Platform for Embodied AI ResearchManolis Savva, Jitendra Malik, Devi Parikh, Dhruv Batra et al.ICCV 2019 · 1,863 citations
- Extreme View SynthesisInchang Choi, Orazio Gallo, Alejandro J. Troccoli, Min H. Kim et al.ICCV 2019 · 207 citations
- Boundless: Generative Adversarial Networks for Image ExtensionDilip Krishnan, Piotr Teterwak, Aaron Sarna, Aaron Maschinot et al.ICCV 2019 · 129 citations
- HoloGAN: Unsupervised Learning of 3D Representations From Natural ImagesThu Nguyen-Phuoc, Chuan Li, Lucas Theis, Christian Richardt et al.ICCV 2019 · 98 citations
- Monocular Neural Image Based Rendering With Continuous View ControlJie Song, Xu Chen, Otmar HilligesICCV 2019 · 85 citations
Related papers
- Simple and Effective Synthesis of Indoor 3D ScenesJing Yu Koh, Harsh Agrawal, Dhruv Batra, Richard Tucker et al.AAAI 2023 · 40 citations
- Worldsheet: Wrapping the World in a 3D Sheet for View Synthesis from a Single ImageRonghang Hu, Nikhila Ravi, Alexander C. Berg, Deepak PathakICCV 2021 · 97 citations
- Ponder: Point Cloud Pre-training via Neural RenderingDi Huang, Sida Peng, Tong He, Honghui Yang et al.ICCV 2023 · 55 citations
- MultiDiff: Consistent Novel View Synthesis from a Single ImageNorman Müller, Katja Schwarz, Barbara Rössle, Lorenzo Porzi et al.CVPR 2024 · 14 citations
- Uncertainty-Aware Diffusion-Guided Refinement of 3D ScenesSarosij Bose, Arindam Dutta, Sayak Nag, Junge Zhang et al.ICCV 2025 · 3 citations
