UniDSeg: Unified Cross-Domain 3D Semantic Segmentation via Visual Foundation Models Prior
Yao Wu, Mingwei Xing, Yachao Zhang, Xiaotong Luo, Yuan Xie, Yanyun Qu
Abstract
3D semantic segmentation using an adapting model trained from a source domain with or without accessing unlabeled target-domain data is the fundamental task in computer vision, containing domain adaptation and domain generalization. The essence of simultaneously solving cross-domain tasks is to enhance the generalizability of the encoder. In light of this, we propose a groundbreaking universal method with the help of off-the-shelf Visual Foundation Models (VFMs) to boost the adaptability and generalizability of cross-domain 3D semantic segmentation, dubbed UniDSeg . Our method explores the VFMs prior and how to harness them, aiming to inherit the recognition ability of VFMs. Specifically, this method introduces layer-wise learnable blocks to the VFMs, which hinges on alternately learning two representations during training: (i) Learning visual prompt. The 3D-to-2D transitional prior and task-shared knowledge is captured from the prompt space, and then (ii) Learning deep query. Spatial Tunability is constructed to the representation of distinct instances driven by prompts in the query space. Integrating these representations into a cross-modal learning framework, UniDSeg efficiently mitigates the domain gap between 2D and 3D modalities, achieving unified cross-domain 3D semantic segmentation. Extensive experiments demonstrate the effectiveness of our method across widely recognized tasks and datasets, all achieving superior performance over state-of-the-art methods. Remarkably, UniDSeg achieves 57.5%/54.4% mIoU on “A2D2/sKITTI” for domain adaptive/generalized tasks. Code is available at https://github.com/Barcaaaa/UniDSeg .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3aec926c-8e30-4baf-9e83-ed8cac49dae5Cited by top-tier papers9
- No Object Is an Island: Enhancing 3D Semantic Segmentation Generalization with Diffusion ModelsFan Li, Xuan Wang, Xuanbin Wang, Zhaoxiang Zhang et al.NeurIPS 2025 · 4 citations
- Point-MoE: Large-Scale Multi-Dataset Training with Mixture-of-Experts for 3D Semantic SegmentationXuweiyi Chen, Wentao Zhou, Aruni RoyChowdhury, Zezhou ChengICLR 2026 · 4 citations
- Target Refocusing via Attention Redistribution for Open-Vocabulary Semantic Segmentation: An Explainability PerspectiveJiahao Li, Yang Lu, Yachao Zhang, Yong Xie et al.AAAI 2026 · 3 citations
- UniDxMD: Towards Unified Representation for Cross-Modal Unsupervised Domain Adaptation in 3D Semantic SegmentationZhengyin Liang, Hui Yin, Min Liang, Qianqian Du et al.ICCV 2025 · 2 citations
- PanDA: Unsupervised Domain Adaptation for Multimodal 3D Panoptic Segmentation in Autonomous DrivingYining Pan, Shijie Li, Yuchen Wu, Xulei Yang et al.CVPR 2026 · 1 citation
Builds on29
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- SemanticKITTI: A Dataset for Semantic Scene Understanding of LiDAR SequencesJens Behley, Martin Garbade, Andres Milioto, Jan Quenzel et al.ICCV 2019 · 2,345 citations
- Transfer Learning from Synthetic to Real LiDAR Point Cloud for Semantic SegmentationAoran Xiao, Jiaxing Huang, Dayan Guan, Fangneng Zhan et al.AAAI 2022 · 144 citations
- Generalize then Adapt: Source-Free Domain Adaptive Semantic SegmentationJogendra Nath Kundu, Akshay R. Kulkarni, Amit Singh, Varun Jampani et al.ICCV 2021 · 143 citations
Related papers
- Unlocking 3D Affordance Segmentation with 2D Semantic KnowledgeYu Huang, Zelin Peng, Changsong Wen, Xiaokang Yang et al.CVPR 2026 · 3 citations
- Unleashing the Power of Visual Foundation Models for Generalizable Semantic SegmentationPeiyuan Tang, Xiaodong Zhang, Chunze Yang, Haoran Yuan et al.AAAI 2025 · 3 citations
- AdaCo: Overcoming Visual Foundation Model Noise in 3D Semantic Segmentation via Adaptive Label CorrectionPufan Zou, Shijia Zhao, Weijie Huang, Qiming Xia et al.AAAI 2025
- Generalizable Knowledge Distillation from Vision Foundation Models for Semantic SegmentationChonghua Lv, Dong Zhao, Shuang Wang, Dou Quan et al.CVPR 2026 · 1 citation
- UniVS: Unified and Universal Video Segmentation with Prompts as QueriesMinghan Li, Shuai Li, Xindong Zhang, Lei ZhangCVPR 2024
