Person Search by Text Attribute Query As Zero-Shot Learning
Qi Dong, Xiatian Zhu, Shaogang Gong
摘要
Existing person search methods predominantly assume the availability of at least one-shot imagery sample of the queried person. This assumption is limited in circumstances where only a brief textual (or verbal) description of the target person is available. In this work, we present a deep learning method for attribute text description based person search without any query imagery. Whilst conventional cross-modality matching methods, such as global visual-textual embedding based zero-shot learning and local individual attribute recognition, are functionally applicable, they are limited by several assumptions invalid to person search in deployment scale, data quality, and/or category name semantics. We overcome these issues by formulating an Attribute-Image Hierarchical Matching (AIHM) model. It is able to more reliably match text attribute descriptions with noisy surveillance person images by jointly learning global category-level and local attribute-level textual-visual embedding as well as matching. Extensive evaluations demonstrate the superiority of our AIHM model over a wide variety of state-of-the-art methods on three publicly available attribute labelled surveillance person search benchmarks: Market-1501, DukeMTMC, and PA100K.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Decentralised Learning from Independent Multi-Domain Labels for Person Re-IdentificationGuile Wu, Shaogang GongAAAI 2021 · 被引用 39 次
- ASMR: Learning Attribute-Based Person Search with Adaptive Semantic Margin RegularizerBoseung Jeong, Jicheol Park, Suha KwakICCV 2021 · 被引用 29 次
- Enhanced Visual-Semantic Interaction with Tailored Prompts for Pedestrian Attribute RecognitionJunyi Wu, Yan Huang, Min Gao, Yuzhen Niu 等CVPR 2025
- Inter-Task Association Critic for Cross-Resolution Person Re-IdentificationZhiyi Cheng, Qi Dong, Shaogang Gong, Xiatian ZhuCVPR 2020
相关 Paper
- Cross-modal Co-occurrence Attributes Alignments for Person Search by LanguageKai Niu, Linjiang Huang, Yan Huang, Peng Wang 等ACM MM 2022 · 被引用 35 次
- Cross-Modal Cross-Domain Moment Alignment Network for Person SearchYa Jing, Wei Wang, Liang Wang, Tieniu TanCVPR 2020
- DCEL: Deep Cross-modal Evidential Learning for Text-Based Person RetrievalShenshen Li, Xing Xu, Yang Yang, Fumin Shen 等ACM MM 2023 · 被引用 56 次
- Pose-Guided Multi-Granularity Attention Network for Text-Based Person SearchYa Jing, Chenyang Si, Junbo Wang, Wei Wang 等AAAI 2020 · 被引用 182 次
- Improving Pedestrian Attribute Recognition With Weakly-Supervised Multi-Scale Attribute-Specific LocalizationChufeng Tang, Lu Sheng, Zhaoxiang Zhang, Xiaolin HuICCV 2019 · 被引用 153 次
