Diffusion-Based Source-Biased Model for Single Domain Generalized Object Detection
Han Jiang, Wenfei Yang, Tianzhu Zhang, Yongdong Zhang
Abstract
Single domain generalized object detection aims to train an object detector on a single source domain and generalize it to any unseen domain. Although existing approaches based on data augmentation exhibit promising results, they overlook domain discrepancies across multiple augmented domains, which limits the performance of object detectors. To tackle these problems, we propose a novel diffusionbased framework, termed SDG-DiffDet, to mitigate the impact of domain gaps on object detectors. The proposed SDG-DiffDet consists of a memory-guided diffusion module and a source-guided denoising module. Specifically, in the memory-guided diffusion module, we design feature statistics memories that mine diverse style information from local parts to augment source features. The augmented features further serve as noise in the diffusion process, enabling the model to capture distribution differences between practical domain distributions. In the source-guided denoising module, we design a text-guided condition to facilitate distribution transfer from any unseen distribution to source distribution in the denoising process. By combining these two designs, our proposed SDG-DiffDet effectively models feature augmentation and target-to-source distribution transfer within a unified diffusion framework, thereby enhancing the detection performance on unseen domains. Extensive experiments demonstrate that the proposed SDG-DiffDet achieves state-of-the-art performance across two challenging scenarios.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 94465b7a-246c-46c8-b4d0-7ec0f87827ebCited by top-tier papers1
Ask how each one uses itBuilds on29
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Directly Denoising Diffusion ModelsDan Zhang, Jingjing Wang, Feng LuoICML 2024 · 11,724 citations
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 6,759 citations
Related papers
- Generalized Diffusion Detector: Mining Robust Features from Diffusion Models for Domain-Generalized DetectionBoyong He, Yuxiang Ji, Qianwen Ye, Zhuoyue Tan et al.CVPR 2025
- Object-Aware Domain Generalization for Object DetectionWooju Lee, Dasol Hong, Hyungtae Lim, Hyun MyungAAAI 2024 · 58 citations
- Diffusion Domain Teacher: Diffusion Guided Domain Adaptive Object DetectorBoyong He, Yuxiang Ji, Zhuoyue Tan, Liaoni WuACM MM 2024 · 10 citations
- Simulating Distribution Dynamics: Liquid Temporal Feature Evolution for Single-Domain Generalized Object DetectionZihao Zhang, Yang Li, Aming Wu, Yahong HanAAAI 2026
- PhysAug: A Physical-guided and Frequency-based Data Augmentation for Single-Domain Generalized Object DetectionXiaoran Xu, Jiangang Yang, Wenhui Shi, Siyuan Ding et al.AAAI 2025 · 15 citations
