Semantic Scene Completion with Cleaner Self
Fengyun Wang, Dong Zhang, Hanwang Zhang, Jinhui Tang, Qianru Sun
Abstract
Semantic Scene Completion (SSC) transforms an image of single-view depth and/or RGB 2D pixels into 3D voxels, each of whose semantic labels are predicted. SSC is a well-known ill-posed problem as the prediction model has to "imagine" what is behind the visible surface, which is usually represented by Truncated Signed Distance Function (TSDF). Due to the sensory imperfection of the depth camera, most existing methods based on the noisy TSDF estimated from depth values suffer from 1) incomplete volumetric predictions and 2) confused semantic labels. To this end, we use the ground-truth 3D voxels to generate a perfect visible surface, called TSDF-CAD, and then train a "cleaner" SSC model. As the model is noise-free, it is expected to focus more on the "imagination" of unseen voxels. Then, we propose to distill the intermediate "cleaner" knowledge into another model with noisy TSDF input. In particular, we use the 3D occupancy feature and the semantic relations of the "cleaner self" to supervise the counterparts of the "noisy self" to respectively address the above two incorrect predictions. Experimental results validate that our method improves the noisy counterparts with 3.1% IoU and 2.2% mIoU for measuring scene completion and SSC, and also achieves new state-of-the-art accuracy on the popular NYU dataset. The code is available at https://github.com/fereenwong/CleanerS .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2608af5a-84c7-4813-978e-5e2a1e4e2f0cCited by top-tier papers8
- Learning Comprehensive Representations with Richer Self for Text-to-Image Person Re-IdentificationShuanglin Yan, Neng Dong, Jun Liu, Liyan Zhang et al.ACM MM 2023 · 55 citations
- Voxel Proposal Network via Multi-Frame Knowledge Distillation for Semantic Scene CompletionLubo Wang, Di Lin, Kairui Yang, Ruonan Liu et al.NeurIPS 2024 · 14 citations
- Skip Mamba Diffusion for Monocular 3D Semantic Scene CompletionLi Liang, Naveed Akhtar, Jordan Vice, Xiangrui Kong et al.AAAI 2025 · 10 citations
- SplatSSC: Decoupled Depth-Guided Gaussian Splatting for Semantic Scene CompletionRui Qian, Haozhi Cao, Tianchen Deng, Shenghai Yuan et al.AAAI 2026 · 2 citations
- TGSFormer: Scalable Temporal Gaussian Splatting for Embodied Semantic Scene CompletionRui Qian, Haozhi Cao, Tianchen Deng, TIANXIN HU et al.CVPR 2026 · 2 citations
Builds on20
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar et al.NeurIPS 2021 · 9,661 citations
- Decoupled Knowledge DistillationBorui Zhao, Quan Cui, Renjie Song, Yiyu Qiu et al.CVPR 2022 · 835 citations
- Show, Attend and Distill: Knowledge Distillation via Attention-based Feature MatchingMingi Ji, Byeongho Heo, Sungrae ParkAAAI 2021 · 194 citations
- Unpaired Deep Image Deraining Using Dual Contrastive LearningXiang Chen, Jinshan Pan, Kui Jiang, Yufeng Li et al.CVPR 2022 · 190 citations
- Learning Joint 2D-3D Representations for Depth CompletionYun Chen, Bin Yang, Ming Liang, Raquel UrtasunICCV 2019 · 190 citations
Related papers
- ForkNet: Multi-Branch Volumetric Semantic Completion From a Single Depth ImageYida Wang, David Joseph Tan, Nassir Navab, Federico TombariICCV 2019 · 67 citations
- SCPNet: Semantic Scene Completion on Point CloudZhaoyang Xia, Youquan Liu, Xin Li, Xinge Zhu et al.CVPR 2023
- Attention-Based Multi-Modal Fusion Network for Semantic Scene CompletionSiqi Li, Changqing Zou, Yipeng Li, Xibin Zhao et al.AAAI 2020 · 68 citations
- Cascaded Context Pyramid for Full-Resolution 3D Semantic Scene CompletionPingping Zhang, Wei Liu, Yinjie Lei, Huchuan Lu et al.ICCV 2019 · 79 citations
- Memory-Augmented Re-Completion for 3D Semantic Scene CompletionYu-Wen Tseng, Sheng-Ping Yang, Jhih-Ciang Wu, I-Bin Liao et al.AAAI 2025 · 3 citations
