Qinco2: Vector Compression and Search with Improved Implicit Neural Codebooks
Théophane Vallaeys, Matthew J. Muckley, Jakob Verbeek, Matthijs Douze
Abstract
Vector quantization is a fundamental technique for compression and large-scale nearest neighbor search. For high-accuracy operating points, multi-codebook quantization associates data vectors with one element from each of multiple codebooks. An example is residual quantization (RQ), which iteratively quantizes the residual error of previous steps. Dependencies between the different parts of the code are, however, ignored in RQ, which leads to suboptimal rate-distortion performance. QINCo recently addressed this inefficiency by using a neural network to determine the quantization codebook in RQ based on the vector reconstruction from previous steps. In this paper we introduce QINCo2 which extends and improves QINCo with (i) improved vector encoding using codeword pre-selection and beam-search, (ii) a fast approximate decoder leveraging codeword pairs to establish accurate short-lists for search, and (iii) an optimized training procedure and network architecture. We conduct experiments on four datasets to evaluate QINCo2 for vector compression and billion-scale nearest neighbor search. We obtain outstanding results in both settings, improving the state-of-the-art reconstruction MSE by 34% for 16-byte vector compression on BigANN, and search accuracy by 24% with 8-byte encodings on Deep1M.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2c098c53-1f79-40af-b0be-e981999c85bbCited by top-tier papers3
- RQ-MoE: Residual Quantization via Mixture of Experts for Efficient Input-Dependent Vector CompressionZhengjia Zhong, Shuyan Ke, Zaizhou Lin, Jiaqi Song et al.ICML 2026
- Compiling Code LLMs into Lightweight ExecutablesJieke Shi, Junda He, Zhou Yang, Chengran Yang et al.FSE 2026
- CrossQ: Task-Aligned Cross-Token Conditional Quantization for Late Interaction RetrievalRohit Kumar Salla, Manoj Saravanan, Ramya AmancherlaICML 2026
Builds on6
- Autoregressive Image Generation using Residual QuantizationDoyup Lee, Chiheon Kim, Saehoon Kim, Minsu Cho et al.CVPR 2022 · 184 citations
- Towards image compression with perfect realism at ultra-low bitratesMarlène Careil, Matthew J. Muckley, Jakob Verbeek, Stéphane LathuilièreICLR 2024 · 122 citations
- Online Clustered CodebookChuanxia Zheng, Andrea VedaldiICCV 2023 · 67 citations
- Unsupervised Neural Quantization for Compressed-Domain Similarity SearchStanislav Morozov, Artem BabenkoICCV 2019 · 31 citations
- Residual Quantization with Implicit Neural CodebooksIris A. M. Huijben, Matthijs Douze, Matthew J. Muckley, Ruud van Sloun et al.ICML 2024 · 23 citations
Related papers
- Boosting Deep Vector Quantization with Progressive Distribution TransformationWeikang Wang, Xin Zhou, Jun Liu, Weifeng Zhang et al.KDD 2025
- SAQ: Pushing the Limits of Vector Quantization through Code Adjustment and Dimension SegmentationHui Li, Shiyuan Deng, Xiao Yan, Xiangyu Zhi et al.SIGMOD 2026 · 1 citation
- Permute, Quantize, and Fine-Tune: Efficient Compression of Neural NetworksJulieta Martinez, Jashan Shewakramani, Ting-Wei Liu, Ioan Andrei Barsan et al.CVPR 2021
- Product Quantizer Aware Inverted Index for Scalable Nearest Neighbor SearchHae-Chan Noh, Taeho Kim, Jae-Pil HeoICCV 2021 · 9 citations
- Quantization Meets Projection: A Happy Marriage for Approximate k-Nearest Neighbor SearchMingyu Yang, Liuchang Jing, Wentao Li, Wei WangVLDB 2026
