DOLG: Single-Stage Image Retrieval with Deep Orthogonal Fusion of Local and Global Features
Min Yang, Dongliang He, Miao Fan, Baorong Shi, Xuetong Xue, Fu Li, Errui Ding, Jizhou Huang
摘要
Image Retrieval is a fundamental task of obtaining images similar to the query one from a database. A common image retrieval practice is to firstly retrieve candidate images via similarity search using global image features and then re-rank the candidates by leveraging their local features. Previous learning-based studies mainly focus on either global or local image representation learning to tackle the retrieval task. In this paper, we abandon the two-stage paradigm and seek to design an effective single-stage solution by integrating local and global information inside images into compact image representations. Specifically, we propose a Deep Orthogonal Local and Global (DOLG) information fusion framework for end-to-end image retrieval. It attentively extracts representative local information with multi-atrous convolutions and self-attention at first. Components orthogonal to the global image representation are then extracted from the local information. At last, the orthogonal components are concatenated with the global representation as a complementary, and then aggregation is performed to generate the final representation. The whole framework is end-to-end differentiable and can be trained with image-level labels. Extensive experimental results validate the effectiveness of our solution and show that our model achieves state-of-the-art image retrieval performances on Revisited Oxford and Paris datasets. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper21
- Correlation Verification for Image RetrievalSeongwon Lee, Hongje Seong, Suhyeon Lee, Euntai KimCVPR 2022 · 被引用 79 次
- Global Features are All You Need for Image Retrieval and RerankingShihao Shao, Kaifeng Chen, Arjun Karpur, Qinghua Cui 等ICCV 2023 · 被引用 69 次
- Towards Universal Image Embeddings: A Large-Scale Dataset and Challenge for Generic Image RepresentationsNikolaos-Antonios Ypsilantis, Kaifeng Chen, Bingyi Cao, Mário Lipovský 等ICCV 2023 · 被引用 31 次
- A New Benchmark and Model for Challenging Image Manipulation DetectionZhenfei Zhang, Mingyang Li, Ming-Ching ChangAAAI 2024 · 被引用 18 次
- Let All Be Whitened: Multi-Teacher Distillation for Efficient Visual RetrievalZhe Ma, Jianfeng Dong, Shouling Ji, Zhenguang Liu 等AAAI 2024 · 被引用 14 次
它引用的顶会 Paper3
- Learning With Average Precision: Training Image Retrieval With a Listwise LossJérôme Revaud, Jon Almazán, Rafael S. Rezende, César Roberto de SouzaICCV 2019 · 被引用 424 次
- Feature Projection for Improved Text ClassificationQi Qin, Wenpeng Hu, Bing LiuACL 2020 · 被引用 66 次
- Google Landmarks Dataset v2 - A Large-Scale Benchmark for Instance-Level Recognition and RetrievalTobias Weyand, André Araújo, Bingyi Cao, Jack SimCVPR 2020
相关 Paper
- Coarse-to-Fine: Learning Compact Discriminative Representation for Single-Stage Image RetrievalYunquan Zhu, Xinkai Gao, Bo Ke, Ruizhi Qiao 等ICCV 2023 · 被引用 8 次
- Learning Token-Based Representation for Image RetrievalHui Wu, Min Wang, Wengang Zhou, Yang Hu 等AAAI 2022 · 被引用 26 次
- Learning Deep Local Features with Multiple Dynamic Attentions for Large-Scale Image RetrievalHui Wu, Min Wang, Wengang Zhou, Houqiang LiICCV 2021 · 被引用 26 次
- Learning Super-Features for Image RetrievalPhilippe Weinzaepfel, Thomas Lucas, Diane Larlus, Yannis KalantidisICLR 2022 · 被引用 56 次
- Heterogeneous Feature Fusion and Cross-modal Alignment for Composed Image RetrievalGangjian Zhang, Shikui Wei, Huaxin Pang, Yao ZhaoACM MM 2021 · 被引用 34 次
