DDIT: Semantic Scene Completion via Deformable Deep Implicit Templates
Haoang Li, Jinhu Dong, Binghui Wen, Ming Gao, Tianyu Huang, Yun-Hui Liu, Daniel Cremers
Abstract
Scene reconstructions are often incomplete due to occlusions and limited viewpoints. There have been efforts to use semantic information for scene completion. However, the completed shapes may be rough and imprecise since respective methods rely on 3D convolution and/or lack effective shape constraints. To overcome these limitations, we propose a semantic scene completion method based on deformable deep implicit templates (DDIT). Specifically, we complete each segmented instance in a scene by deforming a template with a latent code. Such a template is expressed by a deep implicit function in the canonical frame. It abstracts the shape prior of a category, and thus can provide constraints on the overall shape of an instance. Latent code controls the deformation of template to guarantee fine details of an instance. For code prediction, we design a neural network that leverages both intra-and inter-instance information. We also introduce an algorithm to transform instances between the world and canonical frames based on geometric constraints and a hierarchical tree. To further improve accuracy, we jointly optimize the latent code and transformation by enforcing the zero-valued isosurface constraint. In addition, we establish a new dataset to solve different problems of existing datasets. Experiments showed that our DDIT outperforms state-of-the-art approaches.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8c745532-2f24-4610-8f21-c00e05df1ca3Cited by top-tier papers4
- Behind the Veil: Enhanced Indoor 3D Scene Reconstruction with Occluded Surfaces CompletionSu Sun, Cheng Zhao, Yuliang Guo, Ruoyu Wang et al.CVPR 2024 · 3 citations
- Weak-to-Strong 3D Object Detection with X-Ray DistillationAlexander Gambashidze, Aleksandr Dadukin, Maksim Golyadkin, Maria Razzhivina et al.CVPR 2024
- RWKV-PCSSC: Exploring RWKV Model for Point Cloud Semantic Scene CompletionWenzhe He, Xiaojun Chen, Wentang Chen, Hongyu Wang et al.ACM MM 2025
- Point-based Instance Completion with Scene ConstraintsWesley Khademi, Fuxin LiICLR 2025
Builds on15
- Fully Convolutional Geometric FeaturesChristopher B. Choy, Jaesik Park, Vladlen KoltunICCV 2019 · 807 citations
- LPD-Net: 3D Point Cloud Learning for Large-Scale Place Recognition and Environment AnalysisZhe Liu, Shunbo Zhou, Chuanzhe Suo, Peng Yin et al.ICCV 2019 · 337 citations
- ASFM-Net: Asymmetrical Siamese Feature Matching Network for Point CompletionYaqi Xia, Yan Xia, Wei Li, Rui Song et al.ACM MM 2021 · 93 citations
- End-to-End CAD Model Retrieval and 9DoF Alignment in 3D ScansArmen Avetisyan, Angela Dai, Matthias NießnerICCV 2019 · 88 citations
- ForkNet: Multi-Branch Volumetric Semantic Completion From a Single Depth ImageYida Wang, David Joseph Tan, Nassir Navab, Federico TombariICCV 2019 · 67 citations
Related papers
- Local Deep Implicit Functions for 3D ShapeKyle Genova, Forrester Cole, Avneesh Sud, Aaron Sarna et al.CVPR 2020
- DTF-Net: Category-Level Pose Estimation and Shape Reconstruction via Deformable Template FieldHaowen Wang, Zhipeng Fan, Zhen Zhao, Zhengping Che et al.ACM MM 2023 · 6 citations
- Self-supervised Learning of Implicit Shape Representation with Dense Correspondence for Deformable ObjectsBaowen Zhang, Jiahe Li, Xiaoming Deng, Yinda Zhang et al.ICCV 2023 · 10 citations
- Topology-Preserving Shape Reconstruction and Registration via Neural Diffeomorphic FlowShanlin Sun, Kun Han, Deying Kong, Hao Tang et al.CVPR 2022 · 37 citations
- Deep Implicit Templates for 3D Shape RepresentationZerong Zheng, Tao Yu, Qionghai Dai, Yebin LiuCVPR 2021
