Privacy Protection in Deep Multi-modal Retrieval
Peng-Fei Zhang, Yang Li, Zi Huang, Hongzhi Yin
Abstract
Deep learning techniques have ushered in significant progress in large-scale multi-modal retrieval. Nevertheless, the advanced techniques may be used nefariously to conduct a search that violates the privacy of individuals. In this paper, we propose a novel PrIvacy Protection method (PIP) against malicious multi-modal retrieval models, which proactively transfers original data into adversarial data with quasi-imperceptible perturbations before releasing them. Consequently, unauthorized malicious parties are not able to use deployed deep models to find out desired sensitive information with them. In addition to privacy preserving, PIP synchronously learns an effective multi-modal retrieval model to facilitate authorized uses, endowed with strong resilience to the perturbations. To the best of our knowledge, it is a very first attempt to consider privacy issues in multi-modal retrieval, and encapsulate both privacy protection against unauthorized retrieval and robust multi-modal learning for authorized uses into a unified framework. This work is conducted in the challenging no-box and unsupervised settings, where neither target malicious models nor supervised information is known. The optimization objective of our versatile PIP is achieved through a two-player game between different components with both the intra- and inter-modality graph alignments and the domain distribution alignment considered. Besides, a high-level similarity matrix is developed to obtain reliable guidance for learning. Empirically, we apply the proposed PIP to hashing based multi-modal retrieval scenarios and prove its effectiveness on a range of benchmarks and tasks.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Cited by top-tier papers6
- Universal Adversarial Perturbations for Vision-Language Pre-trained ModelsPeng-Fei Zhang, Zi Huang, Guangdong BaiSIGIR 2024 · 28 citations
- Robust Contrastive Cross-modal Hashing with Noisy LabelsLongan Wang, Yang Qin, Yuan Sun, Dezhong Peng et al.ACM MM 2024 · 14 citations
- Neighbor-aware Contrastive Disambiguation for Cross-Modal Hashing with Redundant AnnotationsChao Su, Likang Peng, Yuan Sun, Dezhong Peng et al.NeurIPS 2025 · 12 citations
- Déjà Vu Memorization in Vision-Language ModelsBargav Jayaraman, Chuan Guo, Kamalika ChaudhuriNeurIPS 2024 · 4 citations
- ICYM2I: The illusion of multimodal informativeness under missingnessYoung Sang Choi, Vincent Jeanselme, Pierre A. Elias, Shalmali JoshiICLR 2026 · 2 citations
Related papers
- Proactive Privacy-preserving Learning for RetrievalPeng-Fei Zhang, Zi Huang, Xin-Shun XuAAAI 2021 · 14 citations
- Prototype-guided Knowledge Transfer for Federated Unsupervised Cross-modal HashingJingzhi Li, Fengling Li, Lei Zhu, Hui Cui et al.ACM MM 2023 · 33 citations
- Once and for All: Universal Transferable Adversarial Perturbation against Deep Hashing-Based Facial Image RetrievalLong Tang, Dengpan Ye, Yunna Lv, Chuanxi Chen et al.AAAI 2024 · 13 citations
- Evade Deep Image Retrieval by Stashing Private Images in the Hash SpaceYanru Xiao, Cong Wang, Xing GaoCVPR 2020
- ShieldIR: Privacy-Preserving Unsupervised Cross-Domain Image Retrieval via Dual Protection TransformationZixin Tang, Haihui Fan, Jinchao Zhang, Hui Ma et al.ACM MM 2025 · 1 citation
