Product Quantizer Aware Inverted Index for Scalable Nearest Neighbor Search
Hae-Chan Noh, Taeho Kim, Jae-Pil Heo
摘要
The inverted index is one of the most commonly used structures for non-exhaustive nearest neighbor search on large-scale datasets. It allows a significant factor of acceleration by a reduced number of distance computations with only a small fraction of the database. In particular, the inverted index enables the product quantization (PQ) to learn their codewords in the residual vector space. The quantization error of the PQ can be substantially improved in such combination since the residual vector space is much more quantization-friendly thanks to their compact distribution compared to the original data. In this paper, we first raise an unremarked but crucial question; why the inverted index and the product quantizer are optimized separately even though they are closely related? For instance, changes on the inverted index distort the whole residual vector space. To address the raised question, we suggest a joint optimization of the coarse and fine quantizers by substituting the original objective of the coarse quantizer to end-to-end quantization distortion. Moreover, our method is generic and applicable to different combinations of coarse and fine quantizers such as inverted multi-index and optimized PQ.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Disentangled Representation Learning for Unsupervised Neural QuantizationHaechan Noh, Sangeek Hyun, Woojin Jeong, Hanshin Lim 等CVPR 2023
- Efficient Precision and Recall Metrics for Assessing Generative Models using Hubness-aware SamplingYuanbang Liang, Jing Wu, Yu-Kun Lai, Yipeng QinICML 2024
它引用的顶会 Paper1
相关 Paper
- Qinco2: Vector Compression and Search with Improved Implicit Neural CodebooksThéophane Vallaeys, Matthew J. Muckley, Jakob Verbeek, Matthijs DouzeICLR 2025
- Residual Quantization with Implicit Neural CodebooksIris A. M. Huijben, Matthijs Douze, Matthew J. Muckley, Ruud van Sloun 等ICML 2024 · 被引用 23 次
- Differentiable Optimized Product Quantization and BeyondZepu Lu, Defu Lian, Jin Zhang, Zaixi Zhang 等WWW 2023 · 被引用 11 次
- NeuVSA: A Unified and Efficient Accelerator for Neural Vector SearchZiming Yuan, Lei Dai, Wen Li, Jie Zhang 等HPCA 2025 · 被引用 1 次
- Leanor: A Learning-Based Accelerator for Efficient Approximate Nearest Neighbor Search via Reduced Memory AccessYi Wang, Huan Liu, Jianan Yuan, Jiaxian Chen 等DAC 2024 · 被引用 4 次
