GTNet: Generative Transfer Network for Zero-Shot Object Detection
Shizhen Zhao, Changxin Gao, Yuanjie Shao, Lerenhan Li, Changqian Yu, Zhong Ji, Nong Sang
Abstract
We propose a Generative Transfer Network (GTNet) for zero-shot object detection (ZSD). GTNet consists of an Object Detection Module and a Knowledge Transfer Module. The Object Detection Module can learn large-scale seen domain knowledge. The Knowledge Transfer Module leverages a feature synthesizer to generate unseen class features, which are applied to train a new classification layer for the Object Detection Module. In order to synthesize features for each unseen class with both the intra-class variance and the IoU variance, we design an IoU-Aware Generative Adversarial Network (IoUGAN) as the feature synthesizer, which can be easily integrated into GTNet. Specifically, IoUGAN consists of three unit models: Class Feature Generating Unit (CFU), Foreground Feature Generating Unit (FFU), and Background Feature Generating Unit (BFU). CFU generates unseen features with the intra-class variance conditioned on the class semantic embeddings. FFU and BFU add the IoU variance to the results of CFU, yielding class-specific foreground and background features, respectively. We evaluate our method on three public datasets and the results demonstrate that our method performs favorably against the state-of-the-art ZSD approaches.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a80271e0-e2d8-4f27-9405-618a0d4b5e6fCited by top-tier papers15
- CoDet: Co-occurrence Guided Region-Word Alignment for Open-Vocabulary Object DetectionChuofan Ma, Yi Jiang, Xin Wen, Zehuan Yuan et al.NeurIPS 2023 · 88 citations
- Robust Region Feature Synthesizer for Zero-Shot Object DetectionPeiliang Huang, Junwei Han, De Cheng, Dingwen ZhangCVPR 2022 · 50 citations
- Open-Vocabulary One-Stage Detection with Hierarchical Visual-Language Knowledge DistillationZongyang Ma, Guan Luo, Jin Gao, Liang Li et al.CVPR 2022 · 44 citations
- Zero-Shot Aerial Object Detection with Visual Description RegularizationZhengqing Zang, Chenyu Lin, Chenwei Tang, Tao Wang et al.AAAI 2024 · 22 citations
- SeeDS: Semantic Separable Diffusion Synthesizer for Zero-shot Food DetectionPengfei Zhou, Weiqing Min, Yang Zhang, Jiajun Song et al.ACM MM 2023 · 11 citations
Related papers
- Self-Supervised Domain-Aware Generative Network for Generalized Zero-Shot LearningJiamin Wu, Tianzhu Zhang, Zheng-Jun Zha, Jiebo Luo et al.CVPR 2020
- Zero-Shot Object Detection by Semantics-Aware DETR with Adaptive Contrastive LossHuan Liu, Lu Zhang, Jihong Guan, Shuigeng ZhouACM MM 2023 · 6 citations
- SAUI: Scale-Aware Unseen Imagineer for Zero-Shot Object DetectionJiahao Wang, Caixia Yan, Weizhan Zhang, Huan Liu et al.AAAI 2024 · 5 citations
- Primitive Generation and Semantic-Related Alignment for Universal Zero-Shot SegmentationShuting He, Henghui Ding, Wei JiangCVPR 2023
- Don't Even Look Once: Synthesizing Features for Zero-Shot DetectionPengkai Zhu, Hanxiao Wang, Venkatesh SaligramaCVPR 2020
