Semantic-guided Reinforced Region Embedding for Generalized Zero-Shot Learning
Jiannan Ge, Hongtao Xie, Shaobo Min, Yongdong Zhang
摘要
Generalized zero-shot Learning (GZSL) aims to recognize images from either seen or unseen domain, mainly by learning a joint embedding space to associate image features with the corresponding category descriptions. Recent methods have proved that localizing important object regions can effectively bridge the semantic-visual gap. However, these are all based on one-off visual localizers, lacking of interpretability and flexibility. In this paper, we propose a novel Semantic-guided Reinforced Region Embedding (SR2E) network that can localize important objects in the long-term interests to construct semantic-visual embedding space. SR2E consists of Reinforced Region Module (R2M) and Semantic Alignment Module (SAM). First, without the annotated bounding box as supervision, R2M encodes the semantic category guidance into the reward and punishment criteria to teach the localizer serialized region searching. Besides, R2M explores different action spaces during the serialized searching path to avoid local optimal localization, which thereby generates discriminative visual features with less redundancy. Second, SAM preserves the semantic relationship into visual features via semantic-visual alignment and designs a domain detector to alleviate the domain confusion. Experiments on four public benchmarks demonstrate that the proposed SR2E is an effective GZSL method with reinforced embedding space, which obtains averaged 6.1% improvements.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- From Two to One: A New Scene Text Recognizer with Visual Language Modeling NetworkYuxin Wang, Hongtao Xie, Shancheng Fang, Jing Wang 等ICCV 2021 · 被引用 184 次
- Towards Balanced Alignment: Modal-Enhanced Semantic Modeling for Video Moment RetrievalZhihang Liu, Jun Li, Hongtao Xie, Pandeng Li 等AAAI 2024 · 被引用 49 次
- Neighborhood-Adaptive Structure Augmented Metric LearningPandeng Li, Yan Li, Hongtao Xie, Lei ZhangAAAI 2022 · 被引用 29 次
- Learning Aligned Cross-Modal Representation for Generalized Zero-Shot ClassificationZhiyu Fang, Xiaobin Zhu, Chun Yang, Zheng Han 等AAAI 2022 · 被引用 26 次
- Distilled Reverse Attention Network for Open-world Compositional Zero-Shot LearningYun Li, Zhe Liu, Saurav Jha, Lina YaoICCV 2023 · 被引用 23 次
它引用的顶会 Paper6
- AU-assisted Graph Attention Convolutional Network for Micro-Expression RecognitionHong-Xia Xie, Ling Lo, Hong-Han Shuai, Wen-Huang ChengACM MM 2020 · 被引用 189 次
- Filtration and Distillation: Enhancing Region Attention for Fine-Grained Visual CategorizationChuanbin Liu, Hongtao Xie, Zheng-Jun Zha, Lingfeng Ma 等AAAI 2020 · 被引用 179 次
- S2SiamFC: Self-supervised Fully Convolutional Siamese Network for Visual TrackingChon-Hou Sio, Yu-Jen Ma, Hong-Han Shuai, Jun-Cheng Chen 等ACM MM 2020 · 被引用 45 次
- Fine-Grained Generalized Zero-Shot Learning via Dense Attribute-Based AttentionDat Huynh, Ehsan ElhamifarCVPR 2020
- Episode-Based Prototype Generating Network for Zero-Shot LearningYunlong Yu, Zhong Ji, Jungong Han, Zhongfei ZhangCVPR 2020
相关 Paper
- Self-Supervised Domain-Aware Generative Network for Generalized Zero-Shot LearningJiamin Wu, Tianzhu Zhang, Zheng-Jun Zha, Jiebo Luo 等CVPR 2020
- A Variational Autoencoder with Deep Embedding Model for Generalized Zero-Shot LearningPeirong Ma, Xiao HuAAAI 2020 · 被引用 43 次
- Task-Independent Knowledge Makes for Transferable Representations for Generalized Zero-Shot LearningChaoqun Wang, Xuejin Chen, Shaobo Min, Xiaoyan Sun 等AAAI 2021 · 被引用 22 次
- Dual Progressive Prototype Network for Generalized Zero-Shot LearningChaoqun Wang, Shaobo Min, Xuejin Chen, Xiaoyan Sun 等NeurIPS 2021 · 被引用 72 次
- Adaptive and Generative Zero-Shot LearningYu-Ying Chou, Hsuan-Tien Lin, Tyng-Luh LiuICLR 2021 · 被引用 25 次
