Learning to Parameterize Visual Attributes for Open-set Fine-grained Retrieval
Shijie Wang, Jianlong Chang, Haojie Li, Zhihui Wang, Wanli Ouyang, Qi Tian
Abstract
Open-set fine-grained retrieval is an emerging challenging task that allows to retrieve unknown categories beyond the training set. The best solution for handling unknown categories is to represent them using a set of visual attributes learnt from known categories, as widely used in zero-shot learning. Though important, attribute modeling usually requires significant manual annotations and thus is labor-intensive. Therefore, it is worth to investigate how to transform retrieval models trained by image-level supervision from category semantic extraction to attribute modeling. To this end, we propose a novel Visual Attribute Parameterization Network (VAPNet) to learn visual attributes from known categories and parameterize them into the retrieval model, without the involvement of any attribute annotations. In this way, VAPNet could utilize its parameters to parse a set of visual attributes from unknown categories and precisely represent them. Technically, VAPNet explicitly attains some semantics with rich details via making use of local image patches and distills the visual attributes from these discovered semantics. Additionally, it integrates the online refinement of these visual attributes into the training process to iteratively enhance their quality. Simultaneously, VAPNet treats these attributes as supervisory signals to tune the retrieval models, thereby achieving attribute parameterization. Extensive experiments on open-set fine-grained retrieval datasets validate the superior performance of our VAPNet over existing solutions.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 95378a46-682c-4c63-a382-d90c304eefc1Cited by top-tier papers1
Ask how each one uses itBuilds on16
- Learning Intra-Batch Connections for Deep Metric LearningJenny Denise Seidenschwarz, Ismail Elezi, Laura Leal-TaixéICML 2021 · 65 citations
- Deep Relational Metric LearningWenzhao Zheng, Borui Zhang, Jiwen Lu, Jie ZhouICCV 2021 · 53 citations
- A-Net: Learning Attribute-Aware Hash Codes for Large-Scale Fine-Grained Image RetrievalXiu-Shen Wei, Yang Shen, Xuhao Sun, Han-Jia Ye et al.NeurIPS 2021 · 48 citations
- Hypergraph-Induced Semantic Tuplet Loss for Deep Metric LearningJongin Lim, Sangdoo Yun, Seulki Park, Jin Young ChoiCVPR 2022 · 41 citations
- Dynamic Position-aware Network for Fine-grained Image RecognitionShijie Wang, Haojie Li, Zhihui Wang, Wanli OuyangAAAI 2021 · 36 citations
Related papers
- Attribute Propagation Network for Graph Zero-Shot LearningLu Liu, Tianyi Zhou, Guodong Long, Jing Jiang et al.AAAI 2020 · 85 citations
- TransZero: Attribute-Guided Transformer for Zero-Shot LearningShiming Chen, Ziming Hong, Yang Liu, Guo-Sen Xie et al.AAAI 2022 · 185 citations
- Attend and Enrich: Enhanced Visual Prompt for Zero-Shot LearningMan Liu, Huihui Bai, Feng Li, Chunjie Zhang et al.AAAI 2025 · 3 citations
- MSDN: Mutually Semantic Distillation Network for Zero-Shot LearningShiming Chen, Ziming Hong, Guo-Sen Xie, Wenhan Yang et al.CVPR 2022 · 141 citations
- Open-Set Fine-Grained Retrieval via Prompting Vision-Language EvaluatorShijie Wang, Jianlong Chang, Haojie Li, Zhihui Wang et al.CVPR 2023
