Lexically-Accelerated Dense Retrieval
Hrishikesh Kulkarni, Sean MacAvaney, Nazli Goharian, Ophir Frieder
摘要
Retrieval approaches that score documents based on learned dense vectors (i.e., dense retrieval) rather than lexical signals (i.e., conventional retrieval) are increasingly popular. Their ability to identify related documents that do not necessarily contain the same terms as those appearing in the user's query (thereby improving recall) is one of their key advantages. However, to actually achieve these gains, dense retrieval approaches typically require an exhaustive search over the document collection, making them considerably more expensive at query-time than conventional lexical approaches. Several techniques aim to reduce this computational overhead by approximating the results of a full dense retriever. Although these approaches reasonably approximate the top results, they suffer in terms of recall -- one of the key advantages of dense retrieval. We introduce 'LADR' (Lexically-Accelerated Dense Retrieval), a simple-yet-effective approach that improves the efficiency of existing dense retrieval models without compromising on retrieval effectiveness. LADR uses lexical retrieval techniques to seed a dense retrieval exploration that uses a document proximity graph. Through extensive experiments, we find that LADR establishes a new dense retrieval effectiveness-efficiency Pareto frontier among approximate k nearest neighbor techniques. When tuned to take around 8ms per query in retrieval latency on our hardware, LADR consistently achieves both precision and recall that are on par with an exhaustive search on standard benchmarks. Importantly, LADR accomplishes this using only a single CPU -- no hardware accelerators such as GPUs -- which reduces the deployment cost of dense retrieval systems.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Improving Retrieval in Theme-specific Applications using a Corpus Topical TaxonomySeongKu Kang, Shivam Agarwal, Bowen Jin, Dongha Lee 等WWW 2024 · 被引用 16 次
- Threshold-driven Pruning with Segmented Maximum Term Weights for Approximate Cluster-based Sparse RetrievalYifan Qiao, Parker Carlson, Shanxiu He, Yingrui Yang 等EMNLP 2024 · 被引用 7 次
- Breaking the Lens of the Telescope: Online Relevance Estimation over Large Retrieval SetsMandeep Rathee, Venktesh V, Sean MacAvaney, Avishek AnandSIGIR 2025 · 被引用 4 次
- QDER: Query-Specific Document and Entity Representations for Multi-Vector Document Re-RankingShubham Chatterjee, Jeff DaltonSIGIR 2025 · 被引用 3 次
- PLAID-PRF: Pseudo-Relevance Feedback with Centroid-like Tokens in PLAIDXiao Wang, Sean MacAvaney, Craig MacdonaldSIGIR 2026
它引用的顶会 Paper6
- Approximate Nearest Neighbor Negative Contrastive Learning for Dense Text RetrievalLee Xiong, Chenyan Xiong, Ye Li, Kwok-Fung Tang 等ICLR 2021 · 被引用 1,547 次
- Accelerating Large-Scale Inference with Anisotropic Vector QuantizationRuiqi Guo, Philip Sun, Erik Lindgren, Quan Geng 等ICML 2020 · 被引用 539 次
- Efficiently Teaching an Effective Dense Retriever with Balanced Topic Aware SamplingSebastian Hofstätter, Sheng-Chieh Lin, Jheng-Hong Yang, Jimmy Lin 等SIGIR 2021 · 被引用 297 次
- Learning Optimal Tree Models under Beam SearchJingwei Zhuo, Ziru Xu, Wei Dai, Han Zhu 等ICML 2020 · 被引用 72 次
- RetroMAE: Pre-Training Retrieval-oriented Language Models Via Masked Auto-EncoderShitao Xiao, Zheng Liu, Yingxia Shao, Zhao CaoEMNLP 2022 · 被引用 63 次
相关 Paper
- Distribution-Driven Dense Retrieval: Modeling Many-to-One Query-Document RelationshipJunfeng Kang, Rui Li, Qi Liu, Zhenya Huang 等AAAI 2025 · 被引用 2 次
- DReX: Accurate and Scalable Dense Retrieval Acceleration via Algorithmic-Hardware CodesignDerrick Quinn, E. Ezgi Yücel, Martin Prammer, Zhenxing Fan 等ISCA 2025 · 被引用 10 次
- LIDER: An Efficient High-dimensional Learned Index for Large-scale Dense Passage RetrievalYifan Wang, Haodi Ma, Daisy Zhe WangVLDB 2023 · 被引用 17 次
- Hybrid Inverted Index Is a Robust Accelerator for Dense RetrievalPeitian Zhang, Zheng Liu, Shitao Xiao, Zhicheng Dou 等EMNLP 2023 · 被引用 4 次
- Scaling Laws for Embedding Dimension in Information RetrievalJulian Killingback, Mahta Rafiee, Madine Manas, Hamed ZamaniSIGIR 2026
