Qinco2: Vector Compression and Search with Improved Implicit Neural Codebooks
Théophane Vallaeys, Matthew J. Muckley, Jakob Verbeek, Matthijs Douze
摘要
Vector quantization is a fundamental technique for compression and large-scale nearest neighbor search. For high-accuracy operating points, multi-codebook quantization associates data vectors with one element from each of multiple codebooks. An example is residual quantization (RQ), which iteratively quantizes the residual error of previous steps. Dependencies between the different parts of the code are, however, ignored in RQ, which leads to suboptimal rate-distortion performance. QINCo recently addressed this inefficiency by using a neural network to determine the quantization codebook in RQ based on the vector reconstruction from previous steps. In this paper we introduce QINCo2 which extends and improves QINCo with (i) improved vector encoding using codeword pre-selection and beam-search, (ii) a fast approximate decoder leveraging codeword pairs to establish accurate short-lists for search, and (iii) an optimized training procedure and network architecture. We conduct experiments on four datasets to evaluate QINCo2 for vector compression and billion-scale nearest neighbor search. We obtain outstanding results in both settings, improving the state-of-the-art reconstruction MSE by 34% for 16-byte vector compression on BigANN, and search accuracy by 24% with 8-byte encodings on Deep1M.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- RQ-MoE: Residual Quantization via Mixture of Experts for Efficient Input-Dependent Vector CompressionZhengjia Zhong, Shuyan Ke, Zaizhou Lin, Jiaqi Song 等ICML 2026
- Compiling Code LLMs into Lightweight ExecutablesJieke Shi, Junda He, Zhou Yang, Chengran Yang 等FSE 2026
- CrossQ: Task-Aligned Cross-Token Conditional Quantization for Late Interaction RetrievalRohit Kumar Salla, Manoj Saravanan, Ramya AmancherlaICML 2026
它引用的顶会 Paper6
- Autoregressive Image Generation using Residual QuantizationDoyup Lee, Chiheon Kim, Saehoon Kim, Minsu Cho 等CVPR 2022 · 被引用 184 次
- Towards image compression with perfect realism at ultra-low bitratesMarlène Careil, Matthew J. Muckley, Jakob Verbeek, Stéphane LathuilièreICLR 2024 · 被引用 122 次
- Online Clustered CodebookChuanxia Zheng, Andrea VedaldiICCV 2023 · 被引用 67 次
- Unsupervised Neural Quantization for Compressed-Domain Similarity SearchStanislav Morozov, Artem BabenkoICCV 2019 · 被引用 31 次
- Residual Quantization with Implicit Neural CodebooksIris A. M. Huijben, Matthijs Douze, Matthew J. Muckley, Ruud van Sloun 等ICML 2024 · 被引用 23 次
相关 Paper
- Boosting Deep Vector Quantization with Progressive Distribution TransformationWeikang Wang, Xin Zhou, Jun Liu, Weifeng Zhang 等KDD 2025
- SAQ: Pushing the Limits of Vector Quantization through Code Adjustment and Dimension SegmentationHui Li, Shiyuan Deng, Xiao Yan, Xiangyu Zhi 等SIGMOD 2026 · 被引用 1 次
- Permute, Quantize, and Fine-Tune: Efficient Compression of Neural NetworksJulieta Martinez, Jashan Shewakramani, Ting-Wei Liu, Ioan Andrei Barsan 等CVPR 2021
- Product Quantizer Aware Inverted Index for Scalable Nearest Neighbor SearchHae-Chan Noh, Taeho Kim, Jae-Pil HeoICCV 2021 · 被引用 9 次
- Quantization Meets Projection: A Happy Marriage for Approximate k-Nearest Neighbor SearchMingyu Yang, Liuchang Jing, Wentao Li, Wei WangVLDB 2026
