ForkNet: Multi-Branch Volumetric Semantic Completion From a Single Depth Image
Yida Wang, David Joseph Tan, Nassir Navab, Federico Tombari
Abstract
We propose a novel model for 3D semantic completion from a single depth image, based on a single encoder and three separate generators used to reconstruct different geometric and semantic representations of the original and completed scene, all sharing the same latent space. To transfer information between the geometric and semantic branches of the network, we introduce paths between them concatenating features at corresponding network layers. Motivated by the limited amount of training samples from real scenes, an interesting attribute of our architecture is the capacity to supplement the existing dataset by generating a new training dataset with high quality, realistic scenes that even includes occlusion and real noise. We build the new dataset by sampling the features directly from latent space which generates a pair of partial volumetric surface and completed volumetric semantic surface. Moreover, we utilize multiple discriminators to increase the accuracy and realism of the reconstructions. We demonstrate the benefits of our approach on standard benchmarks for the two most common completion tasks: semantic 3D scene completion and 3D object completion.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9b5ad0d7-9380-4ba3-9741-8a1f7d799457Cited by top-tier papers18
- MonoScene: Monocular 3D Semantic Scene CompletionAnh-Quan Cao, Raoul de CharetteCVPR 2022 · 251 citations
- Learning Local Displacements for Point Cloud CompletionYida Wang, David Joseph Tan, Nassir Navab, Federico TombariCVPR 2022 · 58 citations
- Not All Voxels Are Equal: Semantic Scene Completion from the Point-Voxel PerspectiveJiaxiang Tang, Xiaokang Chen, Jingbo Wang, Gang ZengAAAI 2022 · 37 citations
- H2GFormer: Horizontal-to-Global Voxel Transformer for 3D Semantic Scene CompletionYu Wang, Chao TongAAAI 2024 · 33 citations
- FFNet: Frequency Fusion Network for Semantic Scene CompletionXuzhi Wang, Di Lin, Liang WanAAAI 2022 · 28 citations
Related papers
- Semantic Scene Completion with Cleaner SelfFengyun Wang, Dong Zhang, Hanwang Zhang, Jinhui Tang et al.CVPR 2023
- DeepHuman: 3D Human Reconstruction From a Single ImageZerong Zheng, Tao Yu, Yixuan Wei, Qionghai Dai et al.ICCV 2019 · 367 citations
- Indoor Scene Generation from a Collection of Semantic-Segmented Depth ImagesMingjia Yang, Yu-Xiao Guo, Bin Zhou, Xin TongICCV 2021 · 41 citations
- Extend3D: Town-Scale 3D GenerationSeungwoo Yoon, Jinmo Kim, Jaesik ParkCVPR 2026 · 4 citations
- 3D Scene Reconstruction With Multi-Layer Depth and Epipolar TransformersDaeyun Shin, Zhile Ren, Erik B. Sudderth, Charless C. FowlkesICCV 2019 · 67 citations
