Weakly-Supervised Image Hashing through Masked Visual-Semantic Graph-based Reasoning
Lu Jin, Zechao Li, Yonghua Pan, Jinhui Tang
摘要
With the popularization of social websites, many methods have been proposed to explore the noisy tags for weakly-supervised image hashing.The main challenge lies in learning appropriate and sufficient information from those noisy tags. To address this issue, this work proposes a novel Masked visual-semantic Graph-based Reasoning Network, termed as MGRN, to learn joint visual-semantic representations for image hashing. Specifically, for each image, MGRN constructs a relation graph to capture the interactions among its associated tags and performs reasoning with Graph Attention Networks (GAT). MGRN randomly masks out one tag and then make GAT to predict this masked tag. This forces the GAT model to capture the dependence between the image and its associated tags, which can well address the problem of noisy tags. Thus it can capture key tags and visual structures from images to learn well-aligned visual-semantic representations. Finally, the auto-encoders is leveraged to learn hash codes that can preserve the local structure of the joint space. Meanwhile, the joint visual-semantic representations are reconstructed from those hash codes by using a decoder. Experimental results on two widely-used benchmark datasets demonstrate the superiority of the proposed method for image retrieval compared with several state-of-the-art methods.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper2
- Two-pronged Strategy: Lightweight Augmented Graph Network Hashing for Scalable Image RetrievalHui Cui, Lei Zhu, Jingjing Li, Zhiyong Cheng 等ACM MM 2021 · 被引用 16 次
- Conformalized Hierarchical Calibration for Uncertainty-Aware Adaptive HashingJunyu Luo, Jinsheng Huang, Yang Xu, Lutong Zou 等ICLR 2026
相关 Paper
- Self-Supervised Multi-Modal Knowledge Graph Contrastive Hashing for Cross-Modal SearchMeiyu Liang, Junping Du, Zhengyang Liang, Yongwang Xing 等AAAI 2024 · 被引用 24 次
- A-Net: Learning Attribute-Aware Hash Codes for Large-Scale Fine-Grained Image RetrievalXiu-Shen Wei, Yang Shen, Xuhao Sun, Han-Jia Ye 等NeurIPS 2021 · 被引用 48 次
- Webly Supervised Knowledge Embedding Model for Visual ReasoningWenbo Zheng, Lan Yan, Chao Gou, Fei-Yue WangCVPR 2020
- Adaptive Graph Attention Based Discrete Hashing for Incomplete Cross-modal RetrievalShuang Zhang, Yue Wu, Lei Shi, Huilong Jin 等AAAI 2026
- An End-To-End Graph Attention Network Hashing for Cross-Modal RetrievalHuilong Jin, Yingxue Zhang, Lei Shi, Shuang Zhang 等NeurIPS 2024 · 被引用 18 次
