Graph Convolutional Multi-modal Hashing for Flexible Multimedia Retrieval
Xu Lu, Lei Zhu, Li Liu, Liqiang Nie, Huaxiang Zhang
Abstract
Multi-modal hashing makes an important contribution to multimedia retrieval, where a key challenge is to encode heterogeneous modalities into compact hash codes. To solve this dilemma, graph-based multi-modal hashing methods generally define individual affinity matrix of each independent modality and apply linear algorithm for heterogeneous modalities fusion and compact hash learning. Several other methods construct graph Laplacian matrix based on semantic information to help learn discriminative hash code. However, these conventional methods roughly ignore the structural similarity of training set and the complex relations among multi-modal samples, which leads to unsatisfactory complementarity of fused hash codes. More notably, they are faced with two other important problems: huge computing and storage costs caused by graph construction and partial modality feature lost problem when incomplete query sample comes. In this paper, we propose a Flexible Graph Convolutional Multi-modal Hashing (FGCMH) method that adopts GCNs with linear complexity to preserve both the modality-individual and modality-fused structural similarity for discriminative hash learning. Necessarily, accurate multimedia retrieval can be performed on complete and incomplete datasets with our method. Specifically, multiple modality-individual GCNs under semantic guidance are proposed to act on each individual modality independently for intra-modality similarity preserving, then the output representations are fused into a fusion graph with adaptive weighting scheme. Hash GCN and semantic GCN, which share parameters in the first two layers, propagate fusion information and generate hash codes under high-level label space supervision. In the query stage, our method adaptively captures various multi-modal contents in a flexible and robust way, even if partial modality features are lost. Experimental results on three publicly datasets show the flexibility and effectiveness of our proposed method.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get f7e11085-1708-4cef-94cb-d7e70fa69f1eCited by top-tier papers5
- Adaptive Structural Similarity Preserving for Unsupervised Cross Modal HashingLiang Li, Baihua Zheng, Weiwei SunACM MM 2022 · 29 citations
- Distribution-Consistency-Guided Multi-modal HashingJin-Yu Liu, Xian-Ling Mao, Tian-Yi Che, Rong-Cheng TuAAAI 2025 · 8 citations
- TRACI: A Data-centric Approach for Multi-Domain Generalization on GraphsYusheng Zhao, Changhu Wang, Xiao Luo, Junyu Luo et al.AAAI 2025 · 6 citations
- HGOE: Hybrid External and Internal Graph Outlier Exposure for Graph Out-of-Distribution DetectionJunwei He, Qianqian Xu, Yangbangyan Jiang, Zitai Wang et al.ACM MM 2024 · 4 citations
- Clustering-Oriented Generative Attribute Graph ImputationMulin Chen, Bocheng Wang, Jiaxin Zhong, Zongcheng Miao et al.ACM MM 2025
Related papers
- Graph Convolutional Incomplete Multi-modal HashingXiaobo Shen, Yinfan Chen, Shirui Pan, Weiwei Liu et al.ACM MM 2023 · 16 citations
- Graph Convolutional Semi-Supervised Cross-Modal HashingXiaobo Shen, Gaoyao Yu, Yinfan Chen, Xichen Yang et al.ACM MM 2024 · 5 citations
- Local Graph Convolutional Networks for Cross-Modal HashingYudong Chen, Sen Wang, Jianglin Lu, Zhi Chen et al.ACM MM 2021 · 31 citations
- Deep Joint-Semantics Reconstructing Hashing for Large-Scale Unsupervised Cross-Modal RetrievalShupeng Su, Zhisheng Zhong, Chao ZhangICCV 2019 · 261 citations
- Distribution Consistency Guided Hashing for Cross-Modal RetrievalYuan Sun, Kaiming Liu, Yongxiang Li, Zhenwen Ren et al.ACM MM 2024 · 11 citations
