3D Scene Reconstruction With Multi-Layer Depth and Epipolar Transformers
Daeyun Shin, Zhile Ren, Erik B. Sudderth, Charless C. Fowlkes
Abstract
We tackle the problem of automatically reconstructing a complete 3D model of a scene from a single RGB image. This challenging task requires inferring the shape of both visible and occluded surfaces. Our approach utilizes viewer-centered, multi-layer representation of scene geometry adapted from recent methods for single object shape completion. To improve the accuracy of view-centered representations for complex scenes, we introduce a novel "Epipolar Feature Transformer" that transfers convolutional network features from an input view to other virtual camera viewpoints, and thus better covers the 3D scene geometry. Unlike existing approaches that first detect and localize objects in 3D, and then infer object shape using category-specific models, our approach is fully convolutional, end-to-end differentiable, and avoids the resolution and memory limitations of voxel representations. We demonstrate the advantages of multi-layer depth representations and epipolar feature transformers on the reconstruction of a large database of indoor scenes.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d87f48c2-b948-4308-8bd3-61b5c0970ab5Cited by top-tier papers15
- SO-Pose: Exploiting Self-Occlusion for Direct 6D Pose EstimationYan Di, Fabian Manhardt, Gu Wang, Xiangyang Ji et al.ICCV 2021 · 163 citations
- Panoptic 3D Scene Reconstruction From a Single RGB ImageManuel Dahnert, Ji Hou, Matthias Nießner, Angela DaiNeurIPS 2021 · 106 citations
- Human-Aware Object Placement for Visual Environment ReconstructionHongwei Yi, Chun-Hao P. Huang, Dimitrios Tzionas, Muhammed Kocabas et al.CVPR 2022 · 61 citations
- DiffDreamer: Towards Consistent Unsupervised Single-view Scene Extrapolation with Conditional Diffusion ModelsShengqu Cai, Eric Ryan Chan, Songyou Peng, Mohamad Shahbazi et al.ICCV 2023 · 55 citations
- Uni-3D: A Universal Model for Panoptic 3D Scene ReconstructionXiang Zhang, Zeyuan Chen, Fangyin Wei, Zhuowen TuICCV 2023 · 24 citations
Builds on2
- Moulding Humans: Non-Parametric 3D Human Shape Estimation From Single ImagesValentin Gabeur, Jean-Sébastien Franco, Xavier Martin, Cordelia Schmid et al.ICCV 2019 · 140 citations
- X-Section: Cross-Section Prediction for Enhanced RGB-D FusionAndrea Nicastro, Ronald Clark, Stefan LeuteneggerICCV 2019 · 15 citations
Related papers
- Holistic 3D Human and Scene Mesh Estimation From Single View ImagesZhenzhen Weng, Serena YeungCVPR 2021
- Slice3D: Multi-Slice, Occlusion-Revealing, Single View 3D ReconstructionYizhi Wang, Wallace P. Lira, Wenqi Wang, Ali Mahdavi-Amiri et al.CVPR 2024 · 4 citations
- Dual-S3D: Hierarchical Dual-Path Selective SSM-CNN for High-Fidelity Implicit ReconstructionLuoxi Zhang, Pragyan Shrestha, Yu Zhou, Chun Xie et al.ICCV 2025
- Complete 3D Human Reconstruction from a Single Incomplete ImageJunying Wang, Jae Shin Yoon, Tuanfeng Y. Wang, Krishna Kumar Singh et al.CVPR 2023
- RevealNet: Seeing Behind Objects in RGB-D ScansJi Hou, Angela Dai, Matthias NießnerCVPR 2020
