Structural Multiplane Image: Bridging Neural View Synthesis and 3D Reconstruction
Mingfang Zhang, Jinglu Wang, Xiao Li, Yifei Huang, Yoichi Sato, Yan Lu
摘要
The Multiplane Image (MPI), containing a set of frontoparallel RGBα layers, is an effective and efficient representation for view synthesis from sparse inputs. Yet, its fixed structure limits the performance, especially for surfaces imaged at oblique angles. We introduce the Structural MPI (S-MPI), where the plane structure approximates 3D scenes concisely. Conveying RGBα contexts with geometricallyfaithful structures, the S-MPI directly bridges view synthesis and 3D reconstruction. It can not only overcome the critical limitations of MPI, i.e., discretization artifacts from sloped surfaces and abuse of redundant layers, and can also acquire planar 3D reconstruction. Despite the intuition and demand of applying S-MPI, great challenges are introduced, e.g., high-fidelity approximation for both RGBα layers and plane poses, multi-view consistency, non-planar regions modeling, and efficient rendering with intersected planes. Accordingly, we propose a transformer-based network based on a segmentation model [4] . It predicts compact and expressive S-MPI layers with their corresponding masks, poses, and RGBα contexts. Non-planar regions are inclusively handled as a special case in our unified framework. Multi-view consistency is ensured by sharing global proxy embeddings, which encode plane-level features covering the complete 3D scenes with aligned coordinates. Intensive experiments show that our method outperforms both previous state-of-the-art MPI-based view synthesis methods and planar reconstruction methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Slice3D: Multi-Slice, Occlusion-Revealing, Single View 3D ReconstructionYizhi Wang, Wallace P. Lira, Wenqi Wang, Ali Mahdavi-Amiri 等CVPR 2024 · 被引用 4 次
- PoInit-of-View: Poisoning Initialization of Views Transfers Across Multiple 3D Reconstruction SystemsWeijie Wang, Songlong Xing, Zhengyu Zhao, Nicu Sebe 等CVPR 2026 · 被引用 1 次
- DreamStereo: Towards Real-Time Stereo Inpainting for HD VideosYuan Huang, Sijie Zhao, Jing Cheng, Hao Xu 等CVPR 2026 · 被引用 1 次
- Geometry-guided Online 3D Video Synthesis with Multi-View Temporal ConsistencyHyunho Ha, Lei Xiao, Christian Richardt, Thu Nguyen-Phuoc 等CVPR 2025
- Towards In-the-wild 3D Plane Reconstruction from a Single ImageJiachen Liu, Rui Yu, Sili Chen, Sharon X. Huang 等CVPR 2025
它引用的顶会 Paper21
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 被引用 4,089 次
- Per-Pixel Classification is Not All You Need for Semantic SegmentationBowen Cheng, Alexander G. Schwing, Alexander KirillovNeurIPS 2021 · 被引用 2,196 次
- MVSNeRF: Fast Generalizable Radiance Field Reconstruction from Multi-View StereoAnpei Chen, Zexiang Xu, Fuqiang Zhao, Xiaoshuai Zhang 等ICCV 2021 · 被引用 1,024 次
- Direct Voxel Grid Optimization: Super-fast Convergence for Radiance Fields ReconstructionCheng Sun, Min Sun, Hwann-Tzong ChenCVPR 2022 · 被引用 859 次
- Depth-supervised NeRF: Fewer Views and Faster Training for FreeKangle Deng, Andrew Liu, Jun-Yan Zhu, Deva RamananCVPR 2022 · 被引用 756 次
相关 Paper
- Single-View View Synthesis in the Wild with Learned Adaptive Multiplane ImagesYuxuan Han, Ruicheng Wang, Jiaolong YangSIGGRAPH 2022 · 被引用 65 次
- Efficient View Synthesis and 3D-based Multi-Frame Denoising with Multiplane Feature RepresentationsThomas Tanay, Ales Leonardis, Matteo MaggioniCVPR 2023
- Multi-view Pyramid Transformer: Look Coarser to See BroaderGyeongjin Kang, Seungkwon Yang, Seungtae Nam, Younggeun Lee 等CVPR 2026 · 被引用 8 次
- Tiled Multiplane Images for Practical 3D PhotographyNumair Khan, Lei Xiao, Douglas LanmanICCV 2023 · 被引用 15 次
- Single-View View Synthesis With Multiplane ImagesRichard Tucker, Noah SnavelyCVPR 2020
