Multi-label Pattern Image Retrieval via Attention Mechanism Driven Graph Convolutional Network
Ying Li, Hongwei Zhou, Yeyu Yin, Jiaquan Gao
摘要
Pattern images are artificially designed images which are discriminative in aspects of elements, styles, arrangements and so on. Pattern images are widely used in fields like textile, clothing, art, fashion and graphic design. With the growth of image numbers, pattern image retrieval has great potential in commercial applications and industrial production. However, most of existing content-based image retrieval works mainly focus on describing simple attributes with clear conceptual boundaries, which are not suitable for pattern image retrieval. It is difficult to accurately represent and retrieve pattern images which include complex details and multiple elements. Therefore, in this paper, we collect a new pattern image dataset with multiple labels per image for the pattern image retrieval task. To extract discriminative semantic features of multi-label pattern images and construct high-level topology relationships between features, we further propose an Attention Mechanism Driven Graph Convolutional Network (AMD-GCN). Different layers of the multi-semantic attention module activate regions of interest corresponding to multiple labels, respectively. By embedding the learned labels from attention module into the graph convolutional network, which can capture the dependency of labels on the graph manifold, the AMD-GCN builds an end-to-end framework to extract high-level semantic features with label semantics and inner relationships for retrieval. Experiments on the pattern image dataset show that the proposed method highlights the relevant semantic regions of multiple labels, and achieves higher accuracy than state-of-the-art image retrieval methods.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper5
- Towards Multi-Modal Sarcasm Detection via Hierarchical Congruity Modeling with Knowledge EnhancementHui Liu, Wenya Wang, Haoliang LiEMNLP 2022 · 被引用 91 次
- HyP2 Loss: Beyond Hypersphere Metric Space for Multi-label Image RetrievalChengyin Xu, Zenghao Chai, Zhengzhuo Xu, Chun Yuan 等ACM MM 2022 · 被引用 30 次
- Neural Image Popularity Assessment with Retrieval-augmented TransformerLiya Ji, Chan Ho Park, Zhefan Rao, Qifeng ChenACM MM 2023 · 被引用 6 次
- Semi-Open 3D Object Retrieval via Hierarchical Equilibrium on HypergraphYang Xu, Yifan Feng, Jun Zhang, Jun-Hai Yong 等NeurIPS 2024 · 被引用 3 次
- Towards Multimodal Sentiment Analysis via Hierarchical Correlation Modeling with Semantic Distribution ConstraintsQinfu Xu, Yiwei Wei, Chunlei Wu, Leiquan Wang 等AAAI 2025 · 被引用 3 次
相关 Paper
- Tran-GCN: Multi-label Pattern Image Retrieval via Transformer Driven Graph Convolutional NetworkYing Li, Chunming Guan, Rui Cai, Erwan Ye 等ACM MM 2023 · 被引用 3 次
- Visual-Semantic Matching by Exploring High-Order Attention and DistractionYongzhi Li, Duo Zhang, Yadong MuCVPR 2020
- Multi-Label Patent Categorization with Non-Local Attention-Based Graph Convolutional NetworkPingjie Tang, Meng Jiang, Bryan (Ning) Xia, Jed W. Pitera 等AAAI 2020 · 被引用 52 次
- Cross-Modality Attention with Semantic Graph Embedding for Multi-Label ClassificationRenchun You, Zhiyao Guo, Lei Cui, Xiang Long 等AAAI 2020 · 被引用 221 次
- Adaptive Graph Convolutional Network With Attention Graph Clustering for Co-Saliency DetectionKaihua Zhang, Tengpeng Li, Shiwen Shen, Bo Liu 等CVPR 2020
