Domain-RAG: Retrieval-Guided Compositional Image Generation for Cross-Domain Few-Shot Object Detection
Yu Li, Xingyu Qiu, Yuqian Fu, Jie Chen, Tianwen Qian, Xu Zheng, Danda Pani Paudel, Yanwei Fu, Xuanjing Huang, Luc Van Gool, Yu-Gang Jiang
摘要
Cross-Domain Few-Shot Object Detection (CD-FSOD) aims to detect novel objects with only a handful of labeled samples from previously unseen domains. While data augmentation and generative methods have shown promise in few-shot learning, their effectiveness for CD-FSOD remains unclear due to the need for both visual realism and domain alignment. Existing strategies, such as copy-paste augmentation and text-to-image generation, often fail to preserve the correct object category or produce backgrounds coherent with the target domain, making them non-trivial to apply directly to CD-FSOD. To address these challenges, we propose Domain-RAG, a training-free, retrieval-guided compositional image generation framework tailored for CD-FSOD. Domain-RAG consists of three stages: domain-aware background retrieval, domain-guided background generation, and foreground-background composition. Specifically, the input image is first decomposed into foreground and background regions. We then retrieve semantically and stylistically similar images to guide a generative model in synthesizing a new background, conditioned on both the original and retrieved contexts. Finally, the preserved foreground is composed with the newly generated domain-aligned background to form the generated image. Without requiring any additional supervision or training, Domain-RAG produces high-quality, domain-consistent samples across diverse tasks, including CD-FSOD, remote sensing FSOD, and camouflaged FSOD. Extensive experiments show consistent improvements over strong baselines and establish new state-of-the-art results. The source code and instructions are available at https://github.com/LiYu0524/Domain-RAG .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- EgoNight: Towards Egocentric Vision Understanding at Night with a Challenging BenchmarkDeheng Zhang, Yuqian Fu, Runyi Yang, Yang Miao 等ICLR 2026 · 被引用 19 次
- Seg2Any: Open-set Segmentation-Mask-to-Image Generation with Precise Shape and Semantic ControlDanfeng Li, Hui Zhang, Sheng Wang, Jiacheng Li 等NeurIPS 2025 · 被引用 11 次
- A Closer Look at Cross-Domain Few-Shot Object Detection: Fine-Tuning Matters and Parallel Decoder HelpsXuanlong Yu, Youyang Sha, Longfei Liu, Xi Shen 等CVPR 2026 · 被引用 3 次
- Remedying Target-Domain Astigmatism for Cross-Domain Few-Shot Object DetectionYongwei Jiang, Yixiong Zou, Yuhua Li, Ruixuan LiCVPR 2026 · 被引用 3 次
- Retain and Adapt: Auto-Balanced Model Editing for Open-Vocabulary Object Detection under Domain ShiftsZixuan Duan, Fengyuan Lu, Xunzhi Xiang, Wenbin Li 等ICLR 2026
相关 Paper
- StyleProto: Style-Augmented Prototype Learning for Cross-Domain Few-Shot Object DetectionXi Yang, Quantao XieAAAI 2026
- Rethinking the One-shot Object Detection: Cross-Domain Object SearchYupeng Zhang, Shuqi Zheng, Ruize Han, Yuzhong Feng 等ACM MM 2024 · 被引用 1 次
- DRMix: Decomposition-Recomposition Data Augmentation with Diffusion ModelShuo Wang, Zhichuan Wang, Yanmin Chen, Mengyao Zhou 等ACM MM 2025
- Adapting In-Domain Few-Shot Segmentation to New Domains Without Source Domain RetrainingQi Fan, Kaiqi Liu, Nian Liu, Hisham Cholakkal 等ICCV 2025 · 被引用 4 次
- Textual and Visual Guided Task Adaptation for Source-Free Cross-Domain Few-Shot SegmentationJianming Liu, Wenlong Qiu, Haitao WeiACM MM 2025 · 被引用 2 次
