Robust Region Feature Synthesizer for Zero-Shot Object Detection
Peiliang Huang, Junwei Han, De Cheng, Dingwen Zhang
摘要
Zero-shot object detection aims at incorporating class semantic vectors to realize the detection of (both seen and) unseen classes given an unconstrained test image. In this study, we reveal the core challenges in this research area: how to synthesize robust region features (for unseen objects) that are as intra-class diverse and inter-class separable as the real samples, so that strong unseen object detectors can be trained upon them. To address these challenges, we build a novel zero-shot object detection framework that contains an Intra-class Semantic Diverging component and an Inter-class Structure Preserving component. The former is used to realize the one-to-more mapping to obtain diverse visual features from each class semantic vector, preventing miss-classifying the real unseen objects as image backgrounds. While the latter is used to avoid the synthesized features too scattered to mix up the inter-class and foreground-background relationship. To demonstrate the effectiveness of the proposed approach, comprehensive experiments on PASCAL VOC, COCO, and DIOR datasets are conducted. Notably, our approach achieves the new state-of-the-art performance on PASCAL VOC and COCO and it is the first study to carry out zero-shot object detection in remote sensing imagery.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- Zero-Shot Aerial Object Detection with Visual Description RegularizationZhengqing Zang, Chenyu Lin, Chenwei Tang, Tao Wang 等AAAI 2024 · 被引用 22 次
- Knowledge-Enhanced Dual-Stream Zero-Shot Composed Image RetrievalYucheng Suo, Fan Ma, Linchao Zhu, Yi YangCVPR 2024 · 被引用 20 次
- Weakly Supervised Open-Vocabulary Object DetectionJianghang Lin, Yunhang Shen, Bingquan Wang, Shaohui Lin 等AAAI 2024 · 被引用 18 次
- SeeDS: Semantic Separable Diffusion Synthesizer for Zero-shot Food DetectionPengfei Zhou, Weiqing Min, Yang Zhang, Jiajun Song 等ACM MM 2023 · 被引用 11 次
- Meta-ZSDETR: Zero-shot DETR with Meta-learningLu Zhang, Chenbo Zhang, Jiajia Zhao, Jihong Guan 等ICCV 2023 · 被引用 10 次
它引用的顶会 Paper11
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- FREE: Feature Refinement for Generalized Zero-Shot LearningShiming Chen, Wenjie Wang, Beihao Xia, Qinmu Peng 等ICCV 2021 · 被引用 171 次
- GTNet: Generative Transfer Network for Zero-Shot Object DetectionShizhen Zhao, Changxin Gao, Yuanjie Shao, Lerenhan Li 等AAAI 2020 · 被引用 64 次
- Colar: Effective and Efficient Online Action Detection by Consulting ExemplarsLe Yang, Junwei Han, Dingwen ZhangCVPR 2022 · 被引用 55 次
- GaTector: A Unified Framework for Gaze Object PredictionBinglu Wang, Tao Hu, Baoshan Li, Xiaojuan Chen 等CVPR 2022 · 被引用 3 次
相关 Paper
- Don't Even Look Once: Synthesizing Features for Zero-Shot DetectionPengkai Zhu, Hanxiao Wang, Venkatesh SaligramaCVPR 2020
- Zero-Shot Object Detection by Semantics-Aware DETR with Adaptive Contrastive LossHuan Liu, Lu Zhang, Jihong Guan, Shuigeng ZhouACM MM 2023 · 被引用 6 次
- Primitive Generation and Semantic-Related Alignment for Universal Zero-Shot SegmentationShuting He, Henghui Ding, Wei JiangCVPR 2023
- Semantic-Promoted Debiasing and Background Disambiguation for Zero-Shot Instance SegmentationShuting He, Henghui Ding, Wei JiangCVPR 2023
- Context-Aware Zero-Shot RecognitionRuotian Luo, Ning Zhang, Bohyung Han, Linjie YangAAAI 2020 · 被引用 32 次
