Domain-RAG: Retrieval-Guided Compositional Image Generation for Cross-Domain Few-Shot Object Detection
Yu Li, Xingyu Qiu, Yuqian Fu, Jie Chen, Tianwen Qian, Xu Zheng, Danda Pani Paudel, Yanwei Fu, Xuanjing Huang, Luc Van Gool, Yu-Gang Jiang
Abstract
Cross-Domain Few-Shot Object Detection (CD-FSOD) aims to detect novel objects with only a handful of labeled samples from previously unseen domains. While data augmentation and generative methods have shown promise in few-shot learning, their effectiveness for CD-FSOD remains unclear due to the need for both visual realism and domain alignment. Existing strategies, such as copy-paste augmentation and text-to-image generation, often fail to preserve the correct object category or produce backgrounds coherent with the target domain, making them non-trivial to apply directly to CD-FSOD. To address these challenges, we propose Domain-RAG, a training-free, retrieval-guided compositional image generation framework tailored for CD-FSOD. Domain-RAG consists of three stages: domain-aware background retrieval, domain-guided background generation, and foreground-background composition. Specifically, the input image is first decomposed into foreground and background regions. We then retrieve semantically and stylistically similar images to guide a generative model in synthesizing a new background, conditioned on both the original and retrieved contexts. Finally, the preserved foreground is composed with the newly generated domain-aligned background to form the generated image. Without requiring any additional supervision or training, Domain-RAG produces high-quality, domain-consistent samples across diverse tasks, including CD-FSOD, remote sensing FSOD, and camouflaged FSOD. Extensive experiments show consistent improvements over strong baselines and establish new state-of-the-art results. The source code and instructions are available at https://github.com/LiYu0524/Domain-RAG .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f51198be-c75d-419d-b481-2c38e06834a5Cited by top-tier papers5
- EgoNight: Towards Egocentric Vision Understanding at Night with a Challenging BenchmarkDeheng Zhang, Yuqian Fu, Runyi Yang, Yang Miao et al.ICLR 2026 · 19 citations
- Seg2Any: Open-set Segmentation-Mask-to-Image Generation with Precise Shape and Semantic ControlDanfeng Li, Hui Zhang, Sheng Wang, Jiacheng Li et al.NeurIPS 2025 · 11 citations
- A Closer Look at Cross-Domain Few-Shot Object Detection: Fine-Tuning Matters and Parallel Decoder HelpsXuanlong Yu, Youyang Sha, Longfei Liu, Xi Shen et al.CVPR 2026 · 3 citations
- Remedying Target-Domain Astigmatism for Cross-Domain Few-Shot Object DetectionYongwei Jiang, Yixiong Zou, Yuhua Li, Ruixuan LiCVPR 2026 · 3 citations
- Retain and Adapt: Auto-Balanced Model Editing for Open-Vocabulary Object Detection under Domain ShiftsZixuan Duan, Fengyuan Lu, Xunzhi Xiang, Wenbin Li et al.ICLR 2026
Related papers
- StyleProto: Style-Augmented Prototype Learning for Cross-Domain Few-Shot Object DetectionXi Yang, Quantao XieAAAI 2026
- Rethinking the One-shot Object Detection: Cross-Domain Object SearchYupeng Zhang, Shuqi Zheng, Ruize Han, Yuzhong Feng et al.ACM MM 2024 · 1 citation
- DRMix: Decomposition-Recomposition Data Augmentation with Diffusion ModelShuo Wang, Zhichuan Wang, Yanmin Chen, Mengyao Zhou et al.ACM MM 2025
- Adapting In-Domain Few-Shot Segmentation to New Domains Without Source Domain RetrainingQi Fan, Kaiqi Liu, Nian Liu, Hisham Cholakkal et al.ICCV 2025 · 4 citations
- Textual and Visual Guided Task Adaptation for Source-Free Cross-Domain Few-Shot SegmentationJianming Liu, Wenlong Qiu, Haitao WeiACM MM 2025 · 2 citations
