Cross-Dataset Collaborative Learning for Semantic Segmentation in Autonomous Driving
Li Wang, Dong Li, Han Liu, Jinzhang Peng, Lu Tian, Yi Shan
Abstract
Semantic segmentation is an important task for scene understanding in self-driving cars and robotics, which aims to assign dense labels for all pixels in the image. Existing work typically improves semantic segmentation performance by exploring different network architectures on a target dataset. Little attention has been paid to build a unified system by simultaneously learning from multiple datasets due to the inherent distribution shift across different datasets. In this paper, we propose a simple, flexible, and general method for semantic segmentation, termed Cross-Dataset Collaborative Learning (CDCL). Our goal is to train a unified model for improving the performance in each dataset by leveraging information from all the datasets. Specifically, we first introduce a family of Dataset-Aware Blocks (DAB) as the fundamental computing units of the network, which help capture homogeneous convolutional representations and heterogeneous statistics across different datasets. Second, we present a Dataset Alternation Training (DAT) mechanism to facilitate the collaborative optimization procedure. We conduct extensive evaluations on diverse semantic segmentation datasets for autonomous driving. Experiments demonstrate that our method consistently achieves notable improvements over prior single-dataset and cross-dataset training methods without introducing extra FLOPs. Particularly, with the same architecture of PSPNet (ResNet-18), our method outperforms the single-dataset baseline by 5.65%, 6.57%, and 5.79% mIoU on the validation sets of Cityscapes, BDD100K, CamVid, respectively. We also apply CDCL for point cloud 3D semantic segmentation and achieve improved performance, which further validates the superiority and generality of our method. Code and models will be released.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e98c6437-ae59-4fd4-890e-22619ad96019Cited by top-tier papers4
- Towards Large-Scale 3D Representation Learning with Multi-Dataset Point Prompt TrainingXiaoyang Wu, Zhuotao Tian, Xin Wen, Bohao Peng et al.CVPR 2024 · 39 citations
- Towards Multi-Domain Learning for Generalizable Video Anomaly DetectionMyeongAh Cho, Taeoh Kim, Minho Shim, Dongyoon Wee et al.NeurIPS 2024 · 14 citations
- Automated Label Unification for Multi-Dataset Semantic Segmentation with GNNsRong Ma, Jie Chen, Xiangyang Xue, Jian PuNeurIPS 2024 · 3 citations
- LMSeg: Language-guided Multi-dataset SegmentationQiang Zhou, Yuang Liu, Chaohui Yu, Jingliang Li et al.ICLR 2023 · 2 citations
Builds on7
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar et al.NeurIPS 2021 · 9,661 citations
- SemanticKITTI: A Dataset for Semantic Scene Understanding of LiDAR SequencesJens Behley, Martin Garbade, Andres Milioto, Jan Quenzel et al.ICCV 2019 · 2,345 citations
- Segmenter: Transformer for Semantic SegmentationRobin Strudel, Ricardo Garcia, Ivan Laptev, Cordelia SchmidICCV 2021 · 1,898 citations
- Boundary-Aware Feature Propagation for Scene SegmentationHenghui Ding, Xudong Jiang, Ai Qun Liu, Nadia Magnenat-Thalmann et al.ICCV 2019 · 283 citations
Related papers
- Multi-Source Domain Adaptation With Collaborative Learning for Semantic SegmentationJianzhong He, Xu Jia, Shuaijun Chen, Jianzhuang LiuCVPR 2021
- Generalized Semantic Segmentation by Self-Supervised Source Domain Projection and Multi-Level Contrastive LearningLiwei Yang, Xiang Gu, Jian SunAAAI 2023 · 25 citations
- Multi-Space Alignments Towards Universal LiDAR SegmentationYouquan Liu, Lingdong Kong, Xiaoyang Wu, Runnan Chen et al.CVPR 2024
- MultiSiam: Self-supervised Multi-instance Siamese Representation Learning for Autonomous DrivingKai Chen, Lanqing Hong, Hang Xu, Zhenguo Li et al.ICCV 2021 · 60 citations
- Domain generalization of 3D semantic segmentation in autonomous drivingJules Sanchez, Jean-Emmanuel Deschaud, François GouletteICCV 2023 · 37 citations
