Single-View View Synthesis in the Wild with Learned Adaptive Multiplane Images
Yuxuan Han, Ruicheng Wang, Jiaolong Yang
摘要
This paper deals with the challenging task of synthesizing novel views for in-the-wild photographs. Existing methods have shown promising results leveraging monocular depth estimation and color inpainting with layered depth representations. However, these methods still have limited capability to handle scenes with complex 3D geometry. We propose a new method based on the multiplane image (MPI) representation. To accommodate diverse scene layouts in the wild and tackle the difficulty in producing high-dimensional MPI contents, we design a network structure that consists of two novel modules, one for plane depth adjustment and another for depth-aware color prediction. The former adjusts the initial plane positions using the RGBD context feature and an attention mechanism. Given adjusted depth values, the latter predicts the color and density for each plane separately with proper inter-plane interactions achieved via a feature masking strategy. To train our method, we construct large-scale stereo training data using only unconstrained single-view image collections by a simple yet effective warp-back strategy. The experiments on both synthetic and real datasets demonstrate that our trained model works remarkably well and achieves state-of-the-art results.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper34
- Make-It-3D: High-Fidelity 3D Creation from A Single Image with Diffusion PriorJunshu Tang, Tengfei Wang, Bo Zhang, Ting Zhang 等ICCV 2023 · 被引用 405 次
- 3D-aware Image Generation using 2D Diffusion ModelsJianfeng Xiang, Jiaolong Yang, Binbin Huang, Xin TongICCV 2023 · 被引用 82 次
- MonoNeRD: NeRF-like Representations for Monocular 3D Object DetectionJunkai Xu, Liang Peng, Haoran Chen, Hao Li 等ICCV 2023 · 被引用 54 次
- Neural Canvas: Supporting Scenic Design Prototyping by Integrating 3D Sketching and Generative AIYulin Shen, Yifei Shen, Jiawen Cheng, Chutian Jiang 等CHI 2024 · 被引用 41 次
- ViP-NeRF: Visibility Prior for Sparse Input Neural Radiance FieldsNagabhushan Somraj, Rajiv SoundararajanSIGGRAPH 2023 · 被引用 38 次
它引用的顶会 Paper18
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 被引用 2,647 次
- Digging Into Self-Supervised Monocular Depth EstimationClément Godard, Oisin Mac Aodha, Michael Firman, Gabriel J. BrostowICCV 2019 · 被引用 2,416 次
- Free-Form Image Inpainting With Gated ConvolutionJiahui Yu, Zhe Lin, Jimei Yang, Xiaohui Shen 等ICCV 2019 · 被引用 1,990 次
- Focal Frequency Loss for Image Reconstruction and SynthesisLiming Jiang, Bo Dai, Wayne Wu, Chen Change LoyICCV 2021 · 被引用 422 次
- HoloGAN: Unsupervised Learning of 3D Representations From Natural ImagesThu Nguyen-Phuoc, Chuan Li, Lucas Theis, Christian Richardt 等ICCV 2019 · 被引用 98 次
相关 Paper
- Tiled Multiplane Images for Practical 3D PhotographyNumair Khan, Lei Xiao, Douglas LanmanICCV 2023 · 被引用 15 次
- MPI-Flow: Learning Realistic Optical Flow with Multiplane ImagesYingping Liang, Jiaming Liu, Debing Zhang, Ying FuICCV 2023 · 被引用 12 次
- MINE: Towards Continuous Depth MPI with NeRF for Novel View SynthesisJiaxin Li, Zijian Feng, Qi She, Henghui Ding 等ICCV 2021 · 被引用 189 次
- Single-View View Synthesis With Multiplane ImagesRichard Tucker, Noah SnavelyCVPR 2020
- GenWarp: Single Image to Novel Views with Semantic-Preserving Generative WarpingJunyoung Seo, Kazumi Fukuda, Takashi Shibuya, Takuya Narihira 等NeurIPS 2024 · 被引用 81 次
