Personalized Retrieval over Millions of Items
Hemanth Vemuri, Sheshansh Agrawal, Shivam Mittal, Deepak Saini, Akshay Soni, Abhinav V. Sambasivan, Wenhao Lu, Yajun Wang, Mehul Parsana, Purushottam Kar, Manik Varma
Abstract
Personalized retrieval seeks to retrieve items relevant to a user event (e.g. a page visit or a query) that are adapted to the user's personal preferences. For example, two users who happen to perform the same event such as visiting the same product page or asking the same query should receive potentially distinct recommendations adapted to their individual tastes. Personalization is seldom attempted over catalogs of millions of items since the cost of existing personalization routines scale linearly in the number of candidate items. For example, performing two-sided personalized retrieval (with both event and item embeddings personalized to the user) incurs prohibitive storage and compute costs. Instead, it is common to use non-personalized retrieval to obtain a small shortlist of items over which personalized re-ranking can be done quickly. Despite being scalable, this strategy risks losing items uniquely relevant to a user that fail to get shortlisted during non-personalized retrieval. This paper bridges this gap by developing the XPERT algorithm that identifies a form of two-sided personalization that can be scalably implemented over millions of items and hundreds of millions of users. Key to overcoming the computational challenges of personalized retrieval is a novel concept of morph operators that can be used with arbitrary encoder architectures, completely avoids the steep memory overheads of two-sided personalization, provides millisecond-time inference and offers multi-intent retrieval. On multiple public and proprietary datasets, XPERT offered upto 5% superior recall and AUC than state-of-the-art techniques. Code for XPERT is available at https://github.com/personalizedretrieval/xpert.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itBuilds on6
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Approximate Nearest Neighbor Negative Contrastive Learning for Dense Text RetrievalLee Xiong, Chenyan Xiong, Ye Li, Kwok-Fung Tang et al.ICLR 2021 · 1,547 citations
- SiameseXML: Siamese Networks meet Extreme Classifiers with 100M LabelsKunal Dahiya, Ananye Agarwal, Deepak Saini, Gururaj K et al.ICML 2021 · 61 citations
- GalaXC: Graph Neural Networks with Labelwise Attention for Extreme ClassificationDeepak Saini, Arnav Kumar Jain, Kushal Dave, Jian Jiao et al.WWW 2021 · 49 citations
- Octopus: Comprehensive and Elastic User Representation for the Generation of Recommendation CandidatesZheng Liu, Jianxun Lian, Junhan Yang, Defu Lian et al.SIGIR 2020 · 23 citations
Related papers
- Bridging Explicit and Implicit Intent: Unified Interest Generative Method for Joint Search-Recommendation ModelingDongliang Liao, Chenxing Wang, Yawen ZengWWW 2026
- D2K: Turning Historical Data into Retrievable Knowledge for Recommender SystemsJiarui Qin, Weiwen Liu, Weinan Zhang, Yong YuWWW 2025 · 8 citations
- Long Short-Term Session Search: Joint Personalized Reranking and Next Query PredictionQiannan Cheng, Zhaochun Ren, Yujie Lin, Pengjie Ren et al.WWW 2021 · 19 citations
- CPFair: Personalized Consumer and Producer Fairness Re-ranking for Recommender SystemsMohammadmehdi Naghiaei, Hossein A. Rahmani, Yashar DeldjooSIGIR 2022 · 117 citations
- Evoking User Memory: Personalizing LLM via Recollection-Familiarity Adaptive RetrievalYingyi Zhang, Junyi Li, Wenlin Zhang, Pengyue Jia et al.ICLR 2026 · 11 citations
