Semantic Scene Completion with Cleaner Self
Fengyun Wang, Dong Zhang, Hanwang Zhang, Jinhui Tang, Qianru Sun
摘要
Semantic Scene Completion (SSC) transforms an image of single-view depth and/or RGB 2D pixels into 3D voxels, each of whose semantic labels are predicted. SSC is a well-known ill-posed problem as the prediction model has to "imagine" what is behind the visible surface, which is usually represented by Truncated Signed Distance Function (TSDF). Due to the sensory imperfection of the depth camera, most existing methods based on the noisy TSDF estimated from depth values suffer from 1) incomplete volumetric predictions and 2) confused semantic labels. To this end, we use the ground-truth 3D voxels to generate a perfect visible surface, called TSDF-CAD, and then train a "cleaner" SSC model. As the model is noise-free, it is expected to focus more on the "imagination" of unseen voxels. Then, we propose to distill the intermediate "cleaner" knowledge into another model with noisy TSDF input. In particular, we use the 3D occupancy feature and the semantic relations of the "cleaner self" to supervise the counterparts of the "noisy self" to respectively address the above two incorrect predictions. Experimental results validate that our method improves the noisy counterparts with 3.1% IoU and 2.2% mIoU for measuring scene completion and SSC, and also achieves new state-of-the-art accuracy on the popular NYU dataset. The code is available at https://github.com/fereenwong/CleanerS .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Learning Comprehensive Representations with Richer Self for Text-to-Image Person Re-IdentificationShuanglin Yan, Neng Dong, Jun Liu, Liyan Zhang 等ACM MM 2023 · 被引用 55 次
- Voxel Proposal Network via Multi-Frame Knowledge Distillation for Semantic Scene CompletionLubo Wang, Di Lin, Kairui Yang, Ruonan Liu 等NeurIPS 2024 · 被引用 14 次
- Skip Mamba Diffusion for Monocular 3D Semantic Scene CompletionLi Liang, Naveed Akhtar, Jordan Vice, Xiangrui Kong 等AAAI 2025 · 被引用 10 次
- SplatSSC: Decoupled Depth-Guided Gaussian Splatting for Semantic Scene CompletionRui Qian, Haozhi Cao, Tianchen Deng, Shenghai Yuan 等AAAI 2026 · 被引用 2 次
- TGSFormer: Scalable Temporal Gaussian Splatting for Embodied Semantic Scene CompletionRui Qian, Haozhi Cao, Tianchen Deng, TIANXIN HU 等CVPR 2026 · 被引用 2 次
它引用的顶会 Paper20
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar 等NeurIPS 2021 · 被引用 9,661 次
- Decoupled Knowledge DistillationBorui Zhao, Quan Cui, Renjie Song, Yiyu Qiu 等CVPR 2022 · 被引用 835 次
- Show, Attend and Distill: Knowledge Distillation via Attention-based Feature MatchingMingi Ji, Byeongho Heo, Sungrae ParkAAAI 2021 · 被引用 194 次
- Unpaired Deep Image Deraining Using Dual Contrastive LearningXiang Chen, Jinshan Pan, Kui Jiang, Yufeng Li 等CVPR 2022 · 被引用 190 次
- Learning Joint 2D-3D Representations for Depth CompletionYun Chen, Bin Yang, Ming Liang, Raquel UrtasunICCV 2019 · 被引用 190 次
相关 Paper
- ForkNet: Multi-Branch Volumetric Semantic Completion From a Single Depth ImageYida Wang, David Joseph Tan, Nassir Navab, Federico TombariICCV 2019 · 被引用 67 次
- SCPNet: Semantic Scene Completion on Point CloudZhaoyang Xia, Youquan Liu, Xin Li, Xinge Zhu 等CVPR 2023
- Attention-Based Multi-Modal Fusion Network for Semantic Scene CompletionSiqi Li, Changqing Zou, Yipeng Li, Xibin Zhao 等AAAI 2020 · 被引用 68 次
- Cascaded Context Pyramid for Full-Resolution 3D Semantic Scene CompletionPingping Zhang, Wei Liu, Yinjie Lei, Huchuan Lu 等ICCV 2019 · 被引用 79 次
- Memory-Augmented Re-Completion for 3D Semantic Scene CompletionYu-Wen Tseng, Sheng-Ping Yang, Jhih-Ciang Wu, I-Bin Liao 等AAAI 2025 · 被引用 3 次
