LightLT: A Lightweight Representation Quantization Framework for Long-Tail Data
Haoyu Wang, Ruirui Li, Zhengyang Wang, Xianfeng Tang, Danqing Zhang, Monica Xiao Cheng, Bing Yin, Jasha Droppo, Suhang Wang, Jing Gao
摘要
Search tasks require finding items similar to a given query, making it a crucial aspect of various applications. However, storing and computing similarity for millions or billions of item representations can be computationally expensive. To address this, quantization-based hash methods present memory and inference-efficient solutions by converting continuous representations into non-negative integer codes. Despite their advantages, these methods often encounter difficulties in handling long-tail datasets due to imbalanced class distributions.
To address this, we propose LightLT, a lightweight representation quantization framework tailored for long-tail datasets. LightLT produces compact codebooks and discrete IDs, enabling efficient inference by computing distances between query and codewords. Our framework includes innovative designs: 1) Quantization Step: We select the most similar codeword for continuous inputs using the differentiable argmax operation. 2) Double Skip Quantization Connection Module: This module promotes codebook diversity and stability during training. 3) Training Loss: Our comprehensive loss includes class-weighted cross-entropy, center loss, and ranking loss. 4) Model Ensemble: We incorporate a model ensemble step to improve generalization. Theoretical analysis confirms LightLT's low space and inference complexity. Experimental results demonstrate superior performance compared to state-of-the-art baselines in terms of search accuracy, efficiency, and memory usage.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper19
- Decoupling Representation and Classifier for Long-Tailed RecognitionBingyi Kang, Saining Xie, Marcus Rohrbach, Zhicheng Yan 等ICLR 2020 · 被引用 1,496 次
- Model soups: averaging weights of multiple fine-tuned models improves accuracy without increasing inference timeMitchell Wortsman, Gabriel Ilharco, Samir Yitzhak Gadre, Rebecca Roelofs 等ICML 2022 · 被引用 1,464 次
- Long-tail learning via logit adjustmentAditya Krishna Menon, Sadeep Jayasumana, Ankit Singh Rawat, Himanshu Jain 等ICLR 2021 · 被引用 937 次
- Merging Models with Fisher-Weighted AveragingMichael Matena, Colin RaffelNeurIPS 2022 · 被引用 741 次
- Learning With Average Precision: Training Image Retrieval With a Listwise LossJérôme Revaud, Jon Almazán, Rafael S. Rezende, César Roberto de SouzaICCV 2019 · 被引用 424 次
相关 Paper
- Long-Tail HashingYong Chen, Yuqing Hou, Shu Leng, Qing Zhang 等SIGIR 2021 · 被引用 17 次
- Minimizing FLOPs to Learn Efficient Sparse RepresentationsBiswajit Paria, Chih-Kuan Yeh, Ian En-Hsu Yen, Ning Xu 等ICLR 2020 · 被引用 85 次
- Unleashing the Full Potential of Product Quantization for Large-Scale Image RetrievalYu Liang, Shiliang Zhang, Li Ken Li, Xiaoyu WangNeurIPS 2023 · 被引用 5 次
- Efficient Document Retrieval by End-to-End Refining and Quantizing BERT Embedding with Contrastive Product QuantizationZexuan Qiu, Qinliang Su, Jianxing Yu, Shijing SiEMNLP 2022 · 被引用 4 次
- LightRec: A Memory and Search-Efficient Recommender SystemDefu Lian, Haoyu Wang, Zheng Liu, Jianxun Lian 等WWW 2020 · 被引用 106 次
