OA-FSUI2IT: A Novel Few-Shot Cross Domain Object Detection Framework with Object-Aware Few-Shot Unsupervised Image-to-Image Translation
Lifan Zhao, Yunlong Meng, Lin Xu
摘要
Unsupervised image-to-image (UI2I) translation methods aim to learn a mapping between different visual domains with well-preserved content and consistent structure. It has been proven that the generated images are quite useful for enhancing the performance of computer vision tasks like object detection in a different domain with distribution discrepancies. Current methods require large amounts of images in both source and target domains for successful translation. However, data collection and annotations in many scenarios are infeasible or even impossible. In this paper, we propose an Object-Aware Few-Shot UI2I Translation (OA-FSUI2IT) framework to address the few-shot cross domain (FSCD) object detection task with limited unlabeled images in the target domain. To this end, we first introduce a discriminator augmentation (DA) module into the OA-FSUI2IT framework for successful few-shot UI2I translation. Then, we present a patch pyramid contrastive learning (PPCL) strategy to further improve the quality of the generated images. Last, we propose a self-supervised content-consistency (SSCC) loss to enforce the content-consistency in the translation. We implement extensive experiments to demonstrate the effectiveness of our OA-FSUI2IT framework for FSCD object detection and achieve state-of-the-art performance on the benchmarks of Normal-to-Foggy, Day-to-Night, and Cross-scene adaptation. The source code of our proposed method is also available at https://github.com/emdata-ailab/FSCD-Det .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- DON'T NEED RETRAINING: A Mixture of DETR and Vision Foundation Models for Cross-Domain Few-Shot Object DetectionChanghan Liu, Xunzhi Xiang, Zixuan Duan, Wenbin Li 等NeurIPS 2025 · 被引用 8 次
- Open-Scenario Domain Adaptive Object Detection in Autonomous DrivingZeyu Ma, Ziqiang Zheng, Jiwei Wei, Xiaoyong Wei 等ACM MM 2023 · 被引用 2 次
它引用的顶会 Paper8
- Training Generative Adversarial Networks with Limited DataTero Karras, Miika Aittala, Janne Hellsten, Samuli Laine 等NeurIPS 2020 · 被引用 2,345 次
- Differentiable Augmentation for Data-Efficient GAN TrainingShengyu Zhao, Zhijian Liu, Ji Lin, Jun-Yan Zhu 等NeurIPS 2020 · 被引用 707 次
- Few-Shot Unsupervised Image-to-Image TranslationMing-Yu Liu, Xun Huang, Arun Mallya, Tero Karras 等ICCV 2019 · 被引用 668 次
- U-GAT-IT: Unsupervised Generative Attentional Networks with Adaptive Layer-Instance Normalization for Image-to-Image TranslationJunho Kim, Minjae Kim, Hyeonwoo Kang, Kwanghee LeeICLR 2020 · 被引用 632 次
- Reliable Fidelity and Diversity Metrics for Generative ModelsMuhammad Ferjad Naeem, Seong Joon Oh, Youngjung Uh, Yunjey Choi 等ICML 2020 · 被引用 553 次
相关 Paper
- FSCE: Few-Shot Object Detection via Contrastive Proposal EncodingBo Sun, Banghuai Li, Shengcai Cai, Ye Yuan 等CVPR 2021
- AsyFOD: An Asymmetric Adaptation Paradigm for Few-Shot Domain Adaptive Object DetectionYipeng Gao, Kun-Yu Lin, Junkai Yan, Yaowei Wang 等CVPR 2023
- From Dataset to Real-world: General 3D Object Detection via Generalized Cross-domain Few-shot LearningShuangzhi Li, Junlong Shen, Lei Ma, Xingyu LiAAAI 2026
- FiSeR: Fine-Grained Source Representations for Cross-Domain AI Image DetectionShan Zhang, Yongxin He, Mingming Zhang, Huiwen Tian 等ICML 2026
- Semi-Supervised Learning for Few-Shot Image-to-Image TranslationYaxing Wang, Salman H. Khan, Abel Gonzalez-Garcia, Joost van de Weijer 等CVPR 2020
