GTNet: Generative Transfer Network for Zero-Shot Object Detection
Shizhen Zhao, Changxin Gao, Yuanjie Shao, Lerenhan Li, Changqian Yu, Zhong Ji, Nong Sang
摘要
We propose a Generative Transfer Network (GTNet) for zero-shot object detection (ZSD). GTNet consists of an Object Detection Module and a Knowledge Transfer Module. The Object Detection Module can learn large-scale seen domain knowledge. The Knowledge Transfer Module leverages a feature synthesizer to generate unseen class features, which are applied to train a new classification layer for the Object Detection Module. In order to synthesize features for each unseen class with both the intra-class variance and the IoU variance, we design an IoU-Aware Generative Adversarial Network (IoUGAN) as the feature synthesizer, which can be easily integrated into GTNet. Specifically, IoUGAN consists of three unit models: Class Feature Generating Unit (CFU), Foreground Feature Generating Unit (FFU), and Background Feature Generating Unit (BFU). CFU generates unseen features with the intra-class variance conditioned on the class semantic embeddings. FFU and BFU add the IoU variance to the results of CFU, yielding class-specific foreground and background features, respectively. We evaluate our method on three public datasets and the results demonstrate that our method performs favorably against the state-of-the-art ZSD approaches.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper15
- CoDet: Co-occurrence Guided Region-Word Alignment for Open-Vocabulary Object DetectionChuofan Ma, Yi Jiang, Xin Wen, Zehuan Yuan 等NeurIPS 2023 · 被引用 88 次
- Robust Region Feature Synthesizer for Zero-Shot Object DetectionPeiliang Huang, Junwei Han, De Cheng, Dingwen ZhangCVPR 2022 · 被引用 50 次
- Open-Vocabulary One-Stage Detection with Hierarchical Visual-Language Knowledge DistillationZongyang Ma, Guan Luo, Jin Gao, Liang Li 等CVPR 2022 · 被引用 44 次
- Zero-Shot Aerial Object Detection with Visual Description RegularizationZhengqing Zang, Chenyu Lin, Chenwei Tang, Tao Wang 等AAAI 2024 · 被引用 22 次
- SeeDS: Semantic Separable Diffusion Synthesizer for Zero-shot Food DetectionPengfei Zhou, Weiqing Min, Yang Zhang, Jiajun Song 等ACM MM 2023 · 被引用 11 次
相关 Paper
- Self-Supervised Domain-Aware Generative Network for Generalized Zero-Shot LearningJiamin Wu, Tianzhu Zhang, Zheng-Jun Zha, Jiebo Luo 等CVPR 2020
- Zero-Shot Object Detection by Semantics-Aware DETR with Adaptive Contrastive LossHuan Liu, Lu Zhang, Jihong Guan, Shuigeng ZhouACM MM 2023 · 被引用 6 次
- SAUI: Scale-Aware Unseen Imagineer for Zero-Shot Object DetectionJiahao Wang, Caixia Yan, Weizhan Zhang, Huan Liu 等AAAI 2024 · 被引用 5 次
- Primitive Generation and Semantic-Related Alignment for Universal Zero-Shot SegmentationShuting He, Henghui Ding, Wei JiangCVPR 2023
- Don't Even Look Once: Synthesizing Features for Zero-Shot DetectionPengkai Zhu, Hanxiao Wang, Venkatesh SaligramaCVPR 2020
