Diffusion Domain Teacher: Diffusion Guided Domain Adaptive Object Detector
Boyong He, Yuxiang Ji, Zhuoyue Tan, Liaoni Wu
Abstract
Object detectors often suffer a decrease in performance due to the large domain gap between the training data (source domain) and real-world data (target domain). Diffusion-based generative models have shown remarkable abilities in generating high-quality and diverse images, suggesting their potential for extracting valuable feature from various domains. To effectively leverage the crossdomain feature representation of diffusion models, in this paper, we train a detector with frozen-weight diffusion model on the source domain, then employ it as a teacher model to generate pseudo labels on the unlabeled target domain, which are used to guide the supervised learning of the student model on the target domain. We refer to this approach as Diffusion Domain Teacher (DDT). By employing this straightforward yet potent framework, we significantly improve cross-domain object detection performance without compromising the inference speed. Our method achieves an average mAP improvement of 21.2% compared to the baseline on 6 datasets from three common cross-domain detection benchmarks (Cross-Camera, Syn2Real, Real2Artistic), surpassing the current state-of-the-art (SOTA) methods by an average of 5.7% mAP. Furthermore, extensive experiments demonstrate that our method consistently brings improvements even in more powerful and complex models, highlighting broadly applicable and effective domain adaptation capability of our DDT. The code is available at https://github.com/heboyong/Diffusion-Domain-Teacher.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3d1ab26b-3e58-4cc6-8a95-30c46b9545d5Cited by top-tier papers5
- Cloud Object Detector Adaptation by Integrating Different Source KnowledgeShuaifeng Li, Mao Ye, Lihua Zhou, Nianxin Li et al.NeurIPS 2024 · 5 citations
- Boosting Domain Generalized and Adaptive Detection with Diffusion Models: Fitness, Generalization, and TransferabilityBoyong He, Yuxiang Ji, Zhuoyue Tan, Liaoni WuICCV 2025 · 3 citations
- Bridge: Basis-Driven Causal Inference Marries VFMs for Domain GeneralizationMingbo Hong, Feng Liu, Caroline Gevaert, George Vosselman et al.CVPR 2026
- Conditional Diffusion Guided Knowledge Transfer for Multi-Domain Knowledge Graph CompletionJiawei Sheng, Taoyu Su, Xixun Lin, Xiaodong Li et al.WWW 2026
- Generalized Diffusion Detector: Mining Robust Features from Diffusion Models for Domain-Generalized DetectionBoyong He, Yuxiang Ji, Qianwen Ye, Zhuoyue Tan et al.CVPR 2025
Builds on46
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu et al.ICLR 2022 · 18,833 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
Related papers
- Diffusion-Based Source-Biased Model for Single Domain Generalized Object DetectionHan Jiang, Wenfei Yang, Tianzhu Zhang, Yongdong ZhangICCV 2025 · 2 citations
- Contrastive Mean Teacher for Domain Adaptive Object DetectorsShengcao Cao, Dhiraj Joshi, Liang-Yan Gui, Yu-Xiong WangCVPR 2023
- Cross-Domain Adaptive Teacher for Object DetectionYu-Jhe Li, Xiaoliang Dai, Chih-Yao Ma, Yen-Cheng Liu et al.CVPR 2022 · 215 citations
- SimROD: A Simple Adaptation Method for Robust Object DetectionRindra Ramamonjison, Amin Banitalebi-Dehkordi, Xinyu Kang, Xiaolong Bai et al.ICCV 2021 · 66 citations
- Text-Image Alignment for Diffusion-Based PerceptionNeehar Kondapaneni, Markus Marks, Manuel Knott, Rogério Guimarães et al.CVPR 2024
