3D-to-2D Distillation for Indoor Scene Parsing
Zhengzhe Liu, Xiaojuan Qi, Chi-Wing Fu
摘要
Indoor scene semantic parsing from RGB images is very challenging due to occlusions, object distortion, and viewpoint variations. Going beyond prior works that leverage geometry information, typically paired depth maps, we present a new approach, a 3D-to-2D distillation framework, that enables us to leverage 3D features extracted from largescale 3D data repositories (e.g., ScanNet-v2) to enhance 2D features extracted from RGB images. Our work has three novel contributions. First, we distill 3D knowledge from a pretrained 3D network to supervise a 2D network to learn simulated 3D features from 2D features during the training, so the 2D network can infer without requiring 3D data. Second, we design a two-stage dimension normalization scheme to calibrate the 2D and 3D features for better integration. Third, we design a semantic-aware adversarial training model to extend our framework for training with unpaired 3D data. Extensive experiments on various datasets, ScanNet-V2, S3DIS, and NYU-v2, demonstrate the superiority of our approach. Also, experimental results show that our 3D-to-2D distillation improves the model generalization.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper17
- Towards Efficient 3D Object Detection with Knowledge DistillationJihan Yang, Shaoshuai Shi, Runyu Ding, Zhe Wang 等NeurIPS 2022 · 被引用 76 次
- Let Images Give You More: Point Cloud Cross-Modal Training for Shape AnalysisXu Yan, Heshen Zhan, Chaoda Zheng, Jiantao Gao 等NeurIPS 2022 · 被引用 49 次
- EgoDistill: Egocentric Head Motion Distillation for Efficient Video UnderstandingShuhan Tan, Tushar Nagarajan, Kristen GraumanNeurIPS 2023 · 被引用 44 次
- ImLoveNet: Misaligned Image-supported Registration Network for Low-overlap Point Cloud PairsHonghua Chen, Zeyong Wei, Yabin Xu, Mingqiang Wei 等SIGGRAPH 2022 · 被引用 29 次
- Transferring CLIP's Knowledge into Zero-Shot Point Cloud Semantic SegmentationYuanbin Wang, Shaofei Huang, Yulu Gao, Zhen Wang 等ACM MM 2023 · 被引用 17 次
它引用的顶会 Paper3
- KPConv: Flexible and Deformable Convolution for Point CloudsHugues Thomas, Charles R. Qi, Jean-Emmanuel Deschaud, Beatriz Marcotegui 等ICCV 2019 · 被引用 3,193 次
- OccuSeg: Occupancy-Aware 3D Instance SegmentationLei Han, Tian Zheng, Lan Xu, Lu FangCVPR 2020
- PointGroup: Dual-Set Point Grouping for 3D Instance SegmentationLi Jiang, Hengshuang Zhao, Shaoshuai Shi, Shu Liu 等CVPR 2020
相关 Paper
- 3D Segmenter: 3D Transformer based Semantic Segmentation via 2D Panoramic DistillationZhennan Wu, Yang Li, Yifei Huang, Lin Gu 等ICLR 2023
- Pri3D: Can 3D Priors Help 2D Representation Learning?Ji Hou, Saining Xie, Benjamin Graham, Angela Dai 等ICCV 2021 · 被引用 94 次
- Geometry-Aware Network for Domain Adaptive Semantic SegmentationYinghong Liao, Wending Zhou, Xu Yan, Zhen Li 等AAAI 2023 · 被引用 9 次
- GeoGuide: Hierarchical Geometric Guidance for Open-Vocabulary 3D Semantic SegmentationXujing Tao, Chuxin Wang, Yubo Ai, Zhixin Cheng 等CVPR 2026 · 被引用 3 次
- PointDC: Unsupervised Semantic Segmentation of 3D Point Clouds via Cross-modal Distillation and Super-Voxel ClusteringZisheng Chen, Hongbin Xu, Weitao Chen, Zhipeng Zhou 等ICCV 2023 · 被引用 21 次
