Sparse-to-dense Feature Matching: Intra and Inter domain Cross-modal Learning in Domain Adaptation for 3D Semantic Segmentation
Duo Peng, Yinjie Lei, Wen Li, Pingping Zhang, Yulan Guo
摘要
Domain adaptation is critical for success when confronting with the lack of annotations in a new domain. As the huge time consumption of labeling process on 3D point cloud, domain adaptation for 3D semantic segmentation is of great expectation. With the rise of multi-modal datasets, large amount of 2D images are accessible besides 3D point clouds. In light of this, we propose to further leverage 2D data for 3D domain adaptation by intra and inter domain cross modal learning. As for intra-domain cross modal learning, most existing works sample the dense 2D pixel-wise features into the same size with sparse 3D point-wise features, resulting in the abandon of numerous useful 2D features. To address this problem, we propose Dynamic sparse-to-dense Cross Modal Learning (DsCML) to increase the sufficiency of multi-modality information interaction for domain adaptation. For inter-domain cross modal learning, we further advance Cross Modal Adversarial Learning (CMAL) on 2D and 3D data which contains different semantic content aiming to promote high-level modal complementarity. We evaluate our model under various multi-modality domain adaptation settings including day-to-night, country-to-country and dataset-to-dataset, brings large improvements over both uni-modal and multi-modal domain adaptation methods on all settings. Code is available at https://github.com/leolyj/DsCML
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper20
- Segment Any Point Cloud Sequences by Distilling Vision Foundation ModelsYouquan Liu, Lingdong Kong, Jun Cen, Runnan Chen 等NeurIPS 2023 · 被引用 169 次
- Semantic-Aware Domain Generalized SegmentationDuo Peng, Yinjie Lei, Munawar Hayat, Yulan Guo 等CVPR 2022 · 被引用 151 次
- MM-TTA: Multi-Modal Test-Time Adaptation for 3D Semantic SegmentationInkyu Shin, Yi-Hsuan Tsai, Bingbing Zhuang, Samuel Schulter 等CVPR 2022 · 被引用 56 次
- MCD: Diverse Large-Scale Multi-Campus Dataset for Robot PerceptionThien-Minh Nguyen, Shenghai Yuan, Thien Hoang Nguyen, Pengyu Yin 等CVPR 2024 · 被引用 48 次
- Multi-Modal Continual Test-Time Adaptation for 3D Semantic SegmentationHaozhi Cao, Yuecong Xu, Jianfei Yang, Pengyu Yin 等ICCV 2023 · 被引用 27 次
它引用的顶会 Paper4
- SemanticKITTI: A Dataset for Semantic Scene Understanding of LiDAR SequencesJens Behley, Martin Garbade, Andres Milioto, Jan Quenzel 等ICCV 2019 · 被引用 2,345 次
- Confidence Regularized Self-TrainingYang Zou, Zhiding Yu, Xiaofeng Liu, B. V. K. Vijaya Kumar 等ICCV 2019 · 被引用 901 次
- nuScenes: A Multimodal Dataset for Autonomous DrivingHolger Caesar, Varun Bankiti, Alex H. Lang, Sourabh Vora 等CVPR 2020
- RandLA-Net: Efficient Semantic Segmentation of Large-Scale Point CloudsQingyong Hu, Bo Yang, Linhai Xie, Stefano Rosa 等CVPR 2020
相关 Paper
- Cross-Modal Contrastive Learning for Domain Adaptation in 3D Semantic SegmentationBowei Xing, Xianghua Ying, Ruibin Wang, Jinfa Yang 等AAAI 2023 · 被引用 23 次
- Cross-modal & Cross-domain Learning for Unsupervised LiDAR Semantic SegmentationYiyang Chen, Shanshan Zhao, Changxing Ding, Liyao Tang 等ACM MM 2023 · 被引用 4 次
- Self-supervised Exclusive Learning for 3D Segmentation with Cross-Modal Unsupervised Domain AdaptationYachao Zhang, Miaoyu Li, Yuan Xie, Cuihua Li 等ACM MM 2022 · 被引用 22 次
- Mx2M: Masked Cross-Modality Modeling in Domain Adaptation for 3D Semantic SegmentationBoxiang Zhang, Zunran Wang, Yonggen Ling, Yuanyuan Guan 等AAAI 2023 · 被引用 11 次
- Cross-Domain and Cross-Modal Knowledge Distillation in Domain Adaptation for 3D Semantic SegmentationMiaoyu Li, Yachao Zhang, Yuan Xie, Zuodong Gao 等ACM MM 2022 · 被引用 30 次
