Adversarial Attack on Deep Cross-Modal Hamming Retrieval
Chao Li, Shangqian Gao, Cheng Deng, Wei Liu, Heng Huang
Abstract
Recently, Cross-Modal Hamming space Retrieval (CMHR) regains ever-increasing attention, mainly benefiting from the excellent representation capability of deep neural networks. On the other hand, the vulnerability of deep networks exposes a deep cross-modal retrieval system to various safety risks (e.g., adversarial attack). However, attacking deep cross-modal Hamming retrieval remains underexplored. In this paper, we propose an effective Adversarial Attack on Deep Cross-Modal Hamming Retrieval, dubbed AACH, which fools a target deep CMHR model in a black-box setting. Specifically, given a target model, we first construct its substitute model to exploit cross-modal correlations within hamming space, with which we create adversarial examples by limitedly querying from a target model. Furthermore, to enhance the efficiency of adversarial attacks, we design a triplet construction module to exploit cross-modal positive and negative instances. In this way, perturbations can be learned to fool the target model through pulling perturbed examples far away from the positive instances whereas pushing them close to the negative ones. Extensive experiments on three widely used cross-modal (image and text) retrieval benchmarks demonstrate the superiority of the proposed AACH. We find that AACH can successfully attack a given target deep CMHR model with fewer interactions, and that its performance is on par with previous state-of-the-art attacks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 31b28916-df9f-45cb-a298-3a878d571eefCited by top-tier papers4
- VLATTACK: Multimodal Adversarial Attacks on Vision-Language Tasks via Pre-trained ModelsZiyi Yin, Muchao Ye, Tianrong Zhang, Tianyu Du et al.NeurIPS 2023 · 109 citations
- Once and for All: Universal Transferable Adversarial Perturbation against Deep Hashing-Based Facial Image RetrievalLong Tang, Dengpan Ye, Yunna Lv, Chuanxi Chen et al.AAAI 2024 · 13 citations
- HUANG: A Robust Diffusion Model-based Targeted Adversarial Attack Against Deep Hashing RetrievalChihan Huang, Xiaobo ShenAAAI 2025 · 5 citations
- DRFGD: Disentangled Representation-Focused Generative Defense for Attack-Tolerant Cross-Modal HashingZhongqing Yu, Xin Liu, Yiu-ming Cheung, Zhikai Hu et al.AAAI 2026
Builds on5
- Is BERT Really Robust? A Strong Baseline for Natural Language Attack on Text Classification and EntailmentDi Jin, Zhijing Jin, Joey Tianyi Zhou, Peter SzolovitsAAAI 2020 · 1,333 citations
- Deep Joint-Semantics Reconstructing Hashing for Large-Scale Unsupervised Cross-Modal RetrievalShupeng Su, Zhisheng Zhong, Chao ZhangICCV 2019 · 261 citations
- Creating Something From Nothing: Unsupervised Knowledge Distillation for Cross-Modal HashingHengtong Hu, Lingxi Xie, Richang Hong, Qi TianCVPR 2020
- Central Similarity Quantization for Efficient Image and Video RetrievalLi Yuan, Tao Wang, Xiaopeng Zhang, Francis E. H. Tay et al.CVPR 2020
- IMRAM: Iterative Matching With Recurrent Attention Memory for Cross-Modal Image-Text RetrievalHui Chen, Guiguang Ding, Xudong Liu, Zijia Lin et al.CVPR 2020
Related papers
- Vulnerability vs. Reliability: Disentangled Adversarial Examples for Cross-Modal LearningChao Li, Haoteng Tang, Cheng Deng, Liang Zhan et al.KDD 2020 · 19 citations
- You See What I Want You To See: Exploring Targeted Black-Box Transferability Attack for Hash-Based Image Retrieval SystemsYanru Xiao, Cong WangCVPR 2021
- AdvHash: Set-to-set Targeted Attack on Deep Hashing with One Single Adversarial PatchShengshan Hu, Yechao Zhang, Xiaogeng Liu, Leo Yu Zhang et al.ACM MM 2021 · 34 citations
- Prototype-Supervised Adversarial Network for Targeted Attack of Deep HashingXunguang Wang, Zheng Zhang, Baoyuan Wu, Fumin Shen et al.CVPR 2021
- Cross-Modal Stealth: A Coarse-to-Fine Attack Framework for RGB-T TrackerXinyu Xiang, Qinglong Yan, Hao Zhang, Jianfeng Ding et al.AAAI 2025 · 3 citations
