Relevance-Based Embeddings: Lightweight Candidate Retrieval via Heavy-Ranker Calls
Kirill Shevkunov, Andrey Ploskonosov, Liudmila Prokhorenkova
Abstract
In many machine learning applications, the most relevant items for a query should be efficiently retrieved. The relevance function is usually an expensive similarity model, making the exhaustive search infeasible. A typical solution is to train another model that separately embeds queries and items to a vector space, where similarity is defined via the dot product or cosine similarity. This allows one to search the relevant items through fast approximate nearest neighbor search at the cost of some reduction in quality. To compensate for this reduction, the found items (candidates) are re-ranked by the expensive ranking model. In this paper, we investigate an alternative approach to candidate selection that utilizes the scores of the expensive model to improve the representations of queries and items. The idea is to describe each query (item) by its relevance to a set of support items (queries) and use these new representations to obtain query (item) embeddings. We theoretically prove that such embeddings are powerful enough to approximate any complex similarity model (under mild conditions). We also investigate the choice of support items, which is a crucial ingredient of the proposed approach. The experiments on diverse academic and production datasets illustrate the power of our method.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d4be09a9-e6ef-4a05-aec7-1314b8de50d8Builds on8
- Scalable Zero-shot Entity Linking with Dense Entity RetrievalLedell Wu, Fabio Petroni, Martin Josifoski, Sebastian Riedel et al.EMNLP 2020 · 336 citations
- Poly-encoders: Architectures and Pre-training Strategies for Fast and Accurate Multi-sentence ScoringSamuel Humeau, Kurt Shuster, Marie-Anne Lachaux, Jason WestonICLR 2020 · 316 citations
- Dense Passage Retrieval for Open-Domain Question AnsweringVladimir Karpukhin, Barlas Oguz, Sewon Min, Patrick Lewis et al.EMNLP 2020 · 142 citations
- Trans-Encoder: Unsupervised sentence-pair modelling through self- and mutual-distillationsFangyu Liu, Yunlong Jiao, Jordan Massiah, Emine Yilmaz et al.ICLR 2022 · 36 citations
- Efficient Nearest Neighbor Search for Cross-Encoder Models using Matrix FactorizationNishant Yadav, Nicholas Monath, Rico Angell, Manzil Zaheer et al.EMNLP 2022 · 4 citations
Related papers
- Retrieval with Learned SimilaritiesBailu Ding, Jiaqi ZhaiWWW 2025 · 3 citations
- Asymmetric Hashing for Fast Ranking via Neural Network MeasuresKhoa D. Doan, Shulong Tan, Weijie Zhao, Ping LiSIGIR 2023 · 3 citations
- Adaptive Retrieval and Scalable Indexing for k-NN Search with Cross-EncodersNishant Yadav, Nicholas Monath, Manzil Zaheer, Rob Fergus et al.ICLR 2024 · 2 citations
- QSRP: Efficient Reverse k-Ranks Query Processing on High-Dimensional EmbeddingsZheng Bian, Xiao Yan, Jiahao Zhang, Man Lung Yiu et al.ICDE 2024 · 3 citations
- Learning to Select: Query-Aware Adaptive Dimension Selection for Dense RetrievalZhanyu Wu, Richong Zhang, Zhijie NieACL 2026
