Contextual Similarity Aggregation with Self-attention for Visual Re-ranking
Jianbo Ouyang, Hui Wu, Min Wang, Wengang Zhou, Houqiang Li
摘要
In content-based image retrieval, the first-round retrieval result by simple visual feature comparison may be unsatisfactory, which can be refined by visual re-ranking techniques. In image retrieval, it is observed that the contextual similarity among the top-ranked images is an important clue to distinguish the semantic relevance. Inspired by this observation, in this paper, we propose a visual re-ranking method by contextual similarity aggregation with self-attention. In our approach, for each image in the top-K ranking list, we represent it into an affinity feature vector by comparing it with a set of anchor images. Then, the affinity features of the top-K images are refined by aggregating the contextual information with a transformer encoder. Finally, the affinity features are used to recalculate the similarity scores between the query and the top-K images for re-ranking of the latter. To further improve the robustness of our re-ranking model and enhance the performance of our method, a new data augmentation scheme is designed. Since our re-ranking model is not directly involved with the visual feature used in the initial retrieval, it is ready to be applied to retrieval result lists obtained from various retrieval algorithms. We conduct comprehensive experiments on four benchmark datasets to demonstrate the generality and effectiveness of our proposed visual re-ranking method.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Learnable Pillar-based Re-ranking for Image-Text RetrievalLeigang Qu, Meng Liu, Wenjie Wang, Zhedong Zheng 等SIGIR 2023 · 被引用 23 次
- Contextually Affinitive Neighborhood Refinery for Deep ClusteringChunlin Yu, Ye Shi, Jingya WangNeurIPS 2023 · 被引用 15 次
- Cluster-Aware Similarity Diffusion for Instance RetrievalJifei Luo, Hantao Yao, Changsheng XuICML 2024 · 被引用 1 次
- LOCORE: Image Re-ranking with Long-Context Sequence ModelingZilin Xiao, Pavel Suma, Ayush Sachdeva, Hao-Jen Wang 等CVPR 2025
- Locality Preserving Markovian Transition for Instance RetrievalJifei Luo, Wenzheng Wu, Hantao Yao, Lu Yu 等ICML 2025
它引用的顶会 Paper1
相关 Paper
- RAGAR: Retrieval Augmented Personalized Image Generation Guided by RecommendationRun Ling, Wenji Wang, Yuting Liu, Guibing Guo 等AAAI 2026 · 被引用 5 次
- EntRAG: Entity-Centric Retrieval-Augmented Generation for Knowledge-based Visual Question AnsweringYiheng Hu, Xiaoyang Wang, Qing Liu, Sherry Xu 等ICML 2026
- Pix2Key: Controllable Open-Vocabulary Retrieval with Semantic Decomposition and Self-Supervised Visual Dictionary LearningGuoyizhe Wei, Yang Jiao, Nan Xi, Zhishen Huang 等ICML 2026
- Re-Attention for Visual Question AnsweringWenya Guo, Ying Zhang, Xiaoping Wu, Jufeng Yang 等AAAI 2020 · 被引用 90 次
- Correlation Verification for Image RetrievalSeongwon Lee, Hongje Seong, Suhyeon Lee, Euntai KimCVPR 2022 · 被引用 79 次
