Cache-Augmented Inbatch Importance Resampling for Training Recommender Retriever
Jin Chen, Defu Lian, Yucheng Li, Baoyun Wang, Kai Zheng, Enhong Chen
摘要
Recommender retrievers aim to rapidly retrieve a fraction of items from the entire item corpus when a user query requests, with the representative two-tower model trained with the log softmax loss. For efficiently training recommender retrievers on modern hardwares, inbatch sampling, where the items in the mini-batch are shared as negatives to estimate the softmax function, has attained growing interest. However, existing inbatch sampling based strategies just correct the sampling bias of inbatch items with item frequency, being unable to distinguish the user queries within the mini-batch and still incurring significant bias from the softmax. In this paper, we propose a Cache-Augmented Inbatch Importance Resampling ( χ IR) for training recommender retrievers, which not only offers different negatives to user queries with inbatch items, but also adaptively achieves a more accurate estimation of the softmax distribution. Specifically, χ IR resamples items for the given mini-batch training pairs based on certain probabilities, where a cache with more frequently sampled items is adopted to augment the candidate item set, with the purpose of reusing the historical informative samples. χ IR enables to sample query-dependent negatives based on inbatch items and to capture dynamic changes of model training, which leads to a better approximation of the softmax and further contributes to better convergence. Finally, we conduct experiments to validate the superior performance of the proposed χ IR compared with competitive approaches. Preprint. Under review.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Empowering Collaborative Filtering with Principled Adversarial Contrastive LossAn Zhang, Leheng Sheng, Zhibo Cai, Xiang Wang 等NeurIPS 2023 · 被引用 56 次
- CONVERT: Contrastive Graph Clustering with Reliable AugmentationXihong Yang, Cheng Tan, Yue Liu, Ke Liang 等ACM MM 2023 · 被引用 56 次
- Online Distillation-enhanced Multi-modal Transformer for Sequential RecommendationWei Ji, Xiangyan Liu, An Zhang, Yinwei Wei 等ACM MM 2023 · 被引用 32 次
- Preference Diffusion for RecommendationShuo Liu, An Zhang, Guoqing Hu, Hong Qian 等ICLR 2025
它引用的顶会 Paper5
- Accelerating Large-Scale Inference with Anisotropic Vector QuantizationRuiqi Guo, Philip Sun, Erik Lindgren, Quan Geng 等ICML 2020 · 被引用 539 次
- Simplify and Robustify Negative Sampling for Implicit Collaborative FilteringJingtao Ding, Yuhan Quan, Quanming Yao, Yong Li 等NeurIPS 2020 · 被引用 131 次
- Personalized Ranking with Importance SamplingDefu Lian, Qi Liu, Enhong ChenWWW 2020 · 被引用 98 次
- Sampling-Decomposable Generative Adversarial RecommenderBinbin Jin, Defu Lian, Zheng Liu, Qi Liu 等NeurIPS 2020 · 被引用 53 次
- Efficient Training of Retrieval Models using Negative CacheErik Lindgren, Sashank J. Reddi, Ruiqi Guo, Sanjiv KumarNeurIPS 2021 · 被引用 30 次
相关 Paper
- Learning Recommenders for Implicit Feedback with Importance ResamplingJin Chen, Defu Lian, Binbin Jin, Kai Zheng 等WWW 2022 · 被引用 39 次
- Cooperative Retriever and Ranker in Deep RecommendersXu Huang, Defu Lian, Jin Chen, Zheng Liu 等WWW 2023 · 被引用 17 次
- A Fresh Take on Stale Embeddings: Improving Dense Retriever Training with Corrector NetworksNicholas Monath, Will Sussman Grathwohl, Michael Boratko, Rob Fergus 等ICML 2024 · 被引用 1 次
- A Gradient Accumulation Method for Dense Retriever under Memory ConstraintJaehee Kim, Yukyung Lee, Pilsung KangNeurIPS 2024 · 被引用 10 次
- Improving the Accuracy of Dense Retrieval on the Quantized Indexes via Gradient Optimization of the Target EmbeddingsCong Tan, Yongqi Shao, Hong Huo, Tao FangAAAI 2026
