Prototype-guided Knowledge Transfer for Federated Unsupervised Cross-modal Hashing
Jingzhi Li, Fengling Li, Lei Zhu, Hui Cui, Jingjing Li
Abstract
Although deep cross-modal hashing methods have shown superiorities for cross-modal retrieval recently, there is a concern about potential data privacy leakage when training the models. Federated learning adopts a distributed machine learning strategy, which can collaboratively train models without leaking local private data. It is a promising technique to support privacy-preserving cross-modal hashing. However, existing federated learning-based cross-modal retrieval methods usually rely on a large number of semantic annotations, which limits the scalability of the retrieval models. Furthermore, they mostly update the global models by aggregating local model parameters, ignoring the differences in the quantity and category of multi-modal data from multiple clients. To address these issues, we propose a Prototype Transfer-based Federated Unsupervised Cross-modal Hashing(PT-FUCH) method for solving the privacy leakage problem in cross-modal retrieval model learning. PT-FUCH protects local private data by exploring unified global prototypes for different clients, without relying on any semantic annotations. Global prototypes are used to guide the local cross-modal hash learning and promote the alignment of the feature space, thereby alleviating the model bias caused by the difference in the distribution of local multi-modal data and improving the retrieval accuracy. Additionally, we design an adaptive cross-modal knowledge distillation to transfer valuable semantic knowledge from modal-specific global models to local prototype learning processes, reducing the risk of overfitting. Experimental results on three benchmark cross-modal retrieval datasets validate that our PT-FUCH method can achieve outstanding retrieval performance when trained under distributed privacy-preserving mode. The source codes of our method are available at https://github.com/exquisite1210/PT-FUCH_P.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Cited by top-tier papers5
- Asymmetric Cross-Modal Hashing Based on Formal Concept AnalysisYinan Li, Jun Long, Zhan YangAAAI 2025 · 4 citations
- Learning Together Securely: Prototype-Based Federated Multi-Modal Hashing for Safe and Efficient Multi-Modal RetrievalRuifan Zuo, Chaoqun Zheng, Lei Zhu, Wenpeng Lu et al.AAAI 2025 · 4 citations
- Generating Synthetic Data for Unsupervised Federated Learning of Cross-Modal RetrievalTianlong Zhang, Zhe Xue, Adnan Mahmood, Junping Du et al.AAAI 2025 · 3 citations
- Less is More: Federated Graph Learning with Alleviating Topology Heterogeneity from A Causal PerspectiveLele Fu, Bowen Deng, Sheng Huang, Tianchi Liao et al.ICML 2025
- Stationary and Clustering Transformer Hashing for Cross-modal RetrievalZhan Yang, Yiran Liu, Youyuan Huang, Yinan LiAAAI 2026
Related papers
- FedCAFE: Federated Cross-Modal Hashing with Adaptive Feature EnhancementTing Fu, Yu-Wei Zhan, Chong-Yu Zhang, Xin Luo et al.ACM MM 2024 · 5 citations
- Privacy Protection in Deep Multi-modal RetrievalPeng-Fei Zhang, Yang Li, Zi Huang, Hongzhi YinSIGIR 2021 · 21 citations
- HAL: Accurate, Private, and Efficient Sample Alignment for Multimodal Federated LearningXiaokai Zhou, Xiao Yan, Xinyan Li, Yuxiang Wang et al.KDD 2026
- Distribution Consistency Guided Hashing for Cross-Modal RetrievalYuan Sun, Kaiming Liu, Yongxiang Li, Zhenwen Ren et al.ACM MM 2024 · 11 citations
- Graph Convolutional Semi-Supervised Cross-Modal HashingXiaobo Shen, Gaoyao Yu, Yinfan Chen, Xichen Yang et al.ACM MM 2024 · 5 citations
