Single-View View Synthesis in the Wild with Learned Adaptive Multiplane Images
Yuxuan Han, Ruicheng Wang, Jiaolong Yang
Abstract
This paper deals with the challenging task of synthesizing novel views for in-the-wild photographs. Existing methods have shown promising results leveraging monocular depth estimation and color inpainting with layered depth representations. However, these methods still have limited capability to handle scenes with complex 3D geometry. We propose a new method based on the multiplane image (MPI) representation. To accommodate diverse scene layouts in the wild and tackle the difficulty in producing high-dimensional MPI contents, we design a network structure that consists of two novel modules, one for plane depth adjustment and another for depth-aware color prediction. The former adjusts the initial plane positions using the RGBD context feature and an attention mechanism. Given adjusted depth values, the latter predicts the color and density for each plane separately with proper inter-plane interactions achieved via a feature masking strategy. To train our method, we construct large-scale stereo training data using only unconstrained single-view image collections by a simple yet effective warp-back strategy. The experiments on both synthetic and real datasets demonstrate that our trained model works remarkably well and achieves state-of-the-art results.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers34
- Make-It-3D: High-Fidelity 3D Creation from A Single Image with Diffusion PriorJunshu Tang, Tengfei Wang, Bo Zhang, Ting Zhang et al.ICCV 2023 · 405 citations
- 3D-aware Image Generation using 2D Diffusion ModelsJianfeng Xiang, Jiaolong Yang, Binbin Huang, Xin TongICCV 2023 · 82 citations
- MonoNeRD: NeRF-like Representations for Monocular 3D Object DetectionJunkai Xu, Liang Peng, Haoran Chen, Hao Li et al.ICCV 2023 · 54 citations
- Neural Canvas: Supporting Scenic Design Prototyping by Integrating 3D Sketching and Generative AIYulin Shen, Yifei Shen, Jiawen Cheng, Chutian Jiang et al.CHI 2024 · 41 citations
- ViP-NeRF: Visibility Prior for Sparse Input Neural Radiance FieldsNagabhushan Somraj, Rajiv SoundararajanSIGGRAPH 2023 · 38 citations
Builds on18
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 2,647 citations
- Digging Into Self-Supervised Monocular Depth EstimationClément Godard, Oisin Mac Aodha, Michael Firman, Gabriel J. BrostowICCV 2019 · 2,416 citations
- Free-Form Image Inpainting With Gated ConvolutionJiahui Yu, Zhe Lin, Jimei Yang, Xiaohui Shen et al.ICCV 2019 · 1,990 citations
- Focal Frequency Loss for Image Reconstruction and SynthesisLiming Jiang, Bo Dai, Wayne Wu, Chen Change LoyICCV 2021 · 422 citations
- HoloGAN: Unsupervised Learning of 3D Representations From Natural ImagesThu Nguyen-Phuoc, Chuan Li, Lucas Theis, Christian Richardt et al.ICCV 2019 · 98 citations
Related papers
- Tiled Multiplane Images for Practical 3D PhotographyNumair Khan, Lei Xiao, Douglas LanmanICCV 2023 · 15 citations
- MPI-Flow: Learning Realistic Optical Flow with Multiplane ImagesYingping Liang, Jiaming Liu, Debing Zhang, Ying FuICCV 2023 · 12 citations
- MINE: Towards Continuous Depth MPI with NeRF for Novel View SynthesisJiaxin Li, Zijian Feng, Qi She, Henghui Ding et al.ICCV 2021 · 189 citations
- Single-View View Synthesis With Multiplane ImagesRichard Tucker, Noah SnavelyCVPR 2020
- GenWarp: Single Image to Novel Views with Semantic-Preserving Generative WarpingJunyoung Seo, Kazumi Fukuda, Takashi Shibuya, Takuya Narihira et al.NeurIPS 2024 · 81 citations
