LaMP: When Large Language Models Meet Personalization
Alireza Salemi, Sheshera Mysore, Michael Bendersky, Hamed Zamani
Abstract
This paper highlights the importance of personalization in large language models and introduces the LaMP benchmark -a novel benchmark for training and evaluating language models for producing personalized outputs. LaMP offers a comprehensive evaluation framework with diverse language tasks and multiple entries for each user profile. It consists of seven personalized tasks, spanning three text classification and four text generation tasks. We additionally propose two retrieval augmentation approaches that retrieve personal items from each user profile for personalizing language model outputs. To this aim, we study various retrieval models, including term matching, semantic matching, and time-aware methods. Extensive experiments on LaMP for zero-shot and fine-tuned language models demonstrate the efficacy of the proposed retrieval augmentation approach and highlight the impact of personalization in various natural language tasks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 58097fb4-c788-4076-8ef3-60234ce2f617Cited by top-tier papers112
- Rewarded soups: towards Pareto-optimal alignment by interpolating weights fine-tuned on diverse rewardsAlexandre Ramé, Guillaume Couairon, Corentin Dancette, Jean-Baptiste Gaya et al.NeurIPS 2023 · 295 citations
- FLASK: Fine-grained Language Model Evaluation based on Alignment Skill SetsSeonghyeon Ye, Doyoung Kim, Sungdong Kim, Hyeonbin Hwang et al.ICLR 2024 · 176 citations
- HYDRA: Model Factorization Framework for Black-Box LLM PersonalizationYuchen Zhuang, Haotian Sun, Yue Yu, Rushi Qiang et al.NeurIPS 2024 · 79 citations
- Knowledge-Augmented Large Language Models for Personalized Contextual Query SuggestionJinheon Baek, Nirupama Chandrasekaran, Silviu Cucerzan, Allen Herring et al.WWW 2024 · 72 citations
- Large Language Models Empowered Personalized Web AgentsHongru Cai, Yongqi Li, Wenjie Wang, Fengbin Zhu et al.WWW 2025 · 62 citations
Builds on10
- BERTScore: Evaluating Text Generation with BERTTianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger et al.ICLR 2020 · 8,443 citations
- Membership Inference Attacks Against Machine Learning ModelsReza Shokri, Marco Stronati, Congzheng Song, Vitaly ShmatikovS&P 2017 · 5,137 citations
- Jury Learning: Integrating Dissenting Voices into Machine Learning ModelsMitchell L. Gordon, Michelle S. Lam, Joon Sung Park, Kayur Patel et al.CHI 2022 · 134 citations
- Leveraging Similar Users for Personalized Language Modeling with Limited DataCharles Welch, Chenxi Gu, Jonathan K. Kummerfeld, Verónica Pérez-Rosas et al.ACL 2022 · 37 citations
- A Personalized Dense Retrieval Framework for Unified Information AccessHansi Zeng, Surya Kallumadi, Zaid Alibadi, Rodrigo Nogueira et al.SIGIR 2023 · 16 citations
Related papers
- Optimization Methods for Personalizing Large Language Models through Retrieval AugmentationAlireza Salemi, Surya Kallumadi, Hamed ZamaniSIGIR 2024 · 52 citations
- LLMs + Persona-Plug = Personalized LLMsJiongnan Liu, Yutao Zhu, Shuting Wang, Xiaochi Wei et al.ACL 2025 · 19 citations
- LaMP-QA: A Benchmark for Personalized Long-form Question AnsweringAlireza Salemi, Hamed ZamaniEMNLP 2025 · 1 citation
- Retrieval Augmented Generation with Collaborative Filtering for Personalized Text GenerationTeng Shi, Jun Xu, Xiao Zhang, Xiaoxue Zang et al.SIGIR 2025 · 11 citations
- ClusterRAG: Cluster-Based Collaborative Filtering for Personalized Retrieval-Augmented GenerationGibson Nkhata, Uttamasha Anjally Oyshi, Quan Mai, Susan GauchACL 2026
