Inner-Outer Aware Reconstruction Model for Monocular 3D Scene Reconstruction
Yukun Qiu, Guo-Hao Xu, Wei-Shi Zheng
摘要
Monocular 3D scene reconstruction aims to reconstruct the 3D structure of scenes based on posed images. Recent volumetric-based methods directly predict the truncated signed distance function (TSDF) volume and have achieved promising results. The memory cost of volumetric-based methods will grow cubically as the volume size increases, so a coarse-to-fine strategy is necessary for saving memory. Specifically, the coarse-to-fine strategy distinguishes surface voxels from non-surface voxels, and only potential surface voxels are considered in the succeeding procedure. However, the non-surface voxels have various features, and in particular, the voxels on the inner side of the surface are quite different from those on the outer side since there exists an intrinsic gap between them. Therefore, grouping inner-surface and outer-surface voxels into the same class will force the classifier to spend its capacity to bridge the gap. By contrast, it is relatively easy for the classifier to distinguish inner-surface and outer-surface voxels due to the intrinsic gap. Inspired by this, we propose the inner-outer aware reconstruction (IOAR) model. IOAR explores a new coarse-to-fine strategy to classify outer-surface, inner-surface and surface voxels. In addition, IOAR separates occupancy branches from TSDF branches to avoid mutual interference between them. Since our model can better classify the surface, outer-surface and inner-surface voxels, it can predict more precise meshes than existing methods. Experiment results on ScanNet, ICL-NUIM and TUM-RGBD datasets demonstrate the effectiveness and generalization of our model. The code is available at https: //github.com/YorkQiu/InnerOuterAwareReconstruction .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper9
- Neural Sparse Voxel FieldsLingjie Liu, Jiatao Gu, Kyaw Zaw Lin, Tat-Seng Chua 等NeurIPS 2020 · 被引用 1,535 次
- TransformerFusion: Monocular RGB Scene Reconstruction using TransformersAljaz Bozic, Pablo R. Palafox, Justus Thies, Angela Dai 等NeurIPS 2021 · 被引用 185 次
- Multi-View Stereo by Temporal Nonparametric FusionYuxin Hou, Juho Kannala, Arno SolinICCV 2019 · 被引用 99 次
- NeRFusion: Fusing Radiance Fields for Large-Scale Scene ReconstructionXiaoshuai Zhang, Sai Bi, Kalyan Sunkavalli, Hao Su 等CVPR 2022 · 被引用 94 次
- Cascade Cost Volume for High-Resolution Multi-View Stereo and Stereo MatchingXiaodong Gu, Zhiwen Fan, Siyu Zhu, Zuozhuo Dai 等CVPR 2020
相关 Paper
- FineRecon: Depth-aware Feed-forward Network for Detailed 3D ReconstructionNoah Stier, Anurag Ranjan, Alex Colburn, Yajie Yan 等ICCV 2023 · 被引用 29 次
- Monocular Scene Reconstruction with 3D SDF TransformersWeihao Yuan, Xiaodong Gu, Heng Li, Zilong Dong 等ICLR 2023 · 被引用 4 次
- Behind the Veil: Enhanced Indoor 3D Scene Reconstruction with Occluded Surfaces CompletionSu Sun, Cheng Zhao, Yuliang Guo, Ruoyu Wang 等CVPR 2024 · 被引用 3 次
- NeuralRecon: Real-Time Coherent 3D Reconstruction From Monocular VideoJiaming Sun, Yiming Xie, Linghao Chen, Xiaowei Zhou 等CVPR 2021
- MonoNeRD: NeRF-like Representations for Monocular 3D Object DetectionJunkai Xu, Liang Peng, Haoran Chen, Hao Li 等ICCV 2023 · 被引用 54 次
