Sampling Wisely: Deep Image Embedding by Top-K Precision Optimization
Jing Lu, Chaofan Xu, Wei Zhang, Lingyu Duan, Tao Mei
Abstract
Deep image embedding aims at learning a convolutional neural network (CNN) based mapping function that maps an image to a feature vector. The embedding quality is usually evaluated by the performance in image search tasks. Since very few users bother to open the second page search results, top-k precision mostly dominates the user experience and thus is one of the crucial evaluation metrics for the embedding quality. Despite being extensively studied, existing algorithms are usually based on heuristic observation without theoretical guarantee. Consequently, gradient descent direction on the training loss is mostly inconsistent with the direction of optimizing the concerned evaluation metric. This inconsistency certainly misleads the training direction and degrades the performance. In contrast to existing works, in this paper, we propose a novel deep image embedding algorithm with end-to-end optimization to top-k precision, the evaluation metric that is closely related to user experience. Specially, our loss function is constructed with wisely selected ``misplaced" images along the top k nearest neighbor decision boundary, so that the gradient descent update directly promotes the concerned metric, top-k precision. Further more, our theoretical analysis on the upper bounding and consistency properties of the proposed loss supports that minimizing our proposed loss is equivalent to maximizing top-k precision. Experiments show that our proposed algorithm outperforms all compared state-of-the-art deep image embedding algorithms on three benchmark datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 57df4ebe-c978-4a15-b1b7-0a3789396b6bCited by top-tier papers8
- Deep Relational Metric LearningWenzhao Zheng, Borui Zhang, Jiwen Lu, Jie ZhouICCV 2021 · 53 citations
- Recall@k Surrogate Loss with Large Batches and Similarity MixupYash Patel, Giorgos Tolias, Jirí MatasCVPR 2022 · 40 citations
- Distance Metric Learning with Joint Representation DiversificationXu Chu, Yang Lin, Yasha Wang, Xiting Wang et al.ICML 2020 · 11 citations
- SpatialRank: Urban Event Ranking with NDCG Optimization on Spatiotemporal DataBang An, Xun Zhou, Yongjian Zhong, Tianbao YangNeurIPS 2023 · 6 citations
- Talos: Optimizing Top-K Accuracy in Recommender SystemsShengjia Zhang, Weiqin Yang, Jiawei Chen, Peng Wu et al.WWW 2026 · 1 citation
Related papers
- Robust and Decomposable Average Precision for Image RetrievalElias Ramzi, Nicolas Thome, Clément Rambour, Nicolas Audebert et al.NeurIPS 2021 · 40 citations
- Learning With Average Precision: Training Image Retrieval With a Listwise LossJérôme Revaud, Jon Almazán, Rafael S. Rezende, César Roberto de SouzaICCV 2019 · 424 citations
- Deep Metric Learning with Graph ConsistencyBinghui Chen, Pengyu Li, Zhaoyi Yan, Biao Wang et al.AAAI 2021 · 7 citations
- Cross-Image-Attention for Conditional Embeddings in Deep Metric LearningDmytro Kotovenko, Pingchuan Ma, Timo Milbich, Björn OmmerCVPR 2023
- Towards Interpretable Deep Metric Learning with Structural MatchingWenliang Zhao, Yongming Rao, Ziyi Wang, Jiwen Lu et al.ICCV 2021 · 52 citations
