Generative Explore-Exploit: Training-free Optimization of Generative Recommender Systems using LLM Optimizers
Lütfi Kerem Senel, Besnik Fetahu, Davis Yoshida, Zhiyu Chen, Giuseppe Castellucci, Nikhita Vedula, Jason Ingyu Choi, Shervin Malmasi
Abstract
Recommender systems are widely used to suggest engaging content, and Large Language Models (LLMs) have given rise to generative recommenders. Such systems can directly generate items, including for open-set tasks like question suggestion. While the world knowledge of LLMs enable good recommendations, improving the generated content through user feedback is challenging as continuously finetuning LLMs is prohibitively expensive. We present a training-free approach for optimizing generative recommenders by connecting user feedback loops to LLM-based optimizers. We propose a generative explore-exploit method that can not only exploit generated items with known high engagement, but also actively explore and discover hidden population preferences to improve recommendation quality. We evaluate our approach on question generation in two domains (e-commerce and general knowledge), and model user feedback with Click Through Rate (CTR). Experiments show our LLM-based explore-exploit approach can iteratively improve recommendations, and consistently increase CTR. Ablation analysis shows that generative exploration is key to learning user preferences, avoiding the pitfalls of greedy exploit-only approaches. A human evaluation strongly supports our quantitative findings.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2cedf1bf-a00e-4c9b-8c30-cb4e1a60f348Cited by top-tier papers3
- Personal Travel Solver: A Preference-Driven LLM-Solver System for Travel PlanningZijian Shao, Jiancan Wu, Weijian Chen, Xiang WangACL 2025 · 12 citations
- GORACS: Group-level Optimal Transport-guided Coreset Selection for LLM-based Recommender SystemsTiehua Mei, Hengrui Chen, Peng Yu, Jiaqing Liang et al.KDD 2025 · 3 citations
- Reasoning Over Space: Enabling Geographic Reasoning for LLM-Based Generative Next POI RecommendationDongyi Lv, Qiuyu Ding, Heng-Da Xu, Zhaoxu Sun et al.ACL 2026 · 1 citation
Builds on6
- Large Language Models as OptimizersChengrun Yang, Xuezhi Wang, Yifeng Lu, Hanxiao Liu et al.ICLR 2024 · 817 citations
- Recommender Systems with Generative RetrievalShashank Rajput, Nikhil Mehta, Anima Singh, Raghunandan Hulikal Keshavan et al.NeurIPS 2023 · 474 citations
- Rethinking the Evaluation for Conversational Recommendation in the Era of Large Language ModelsXiaolei Wang, Xinyu Tang, Xin Zhao, Jingyuan Wang et al.EMNLP 2023 · 69 citations
- WebCPM: Interactive Web Search for Chinese Long-form Question AnsweringYujia Qin, Zihan Cai, Dian Jin, Lan Yan et al.ACL 2023 · 25 citations
- Generating User-Engaging News HeadlinesPengshan Cai, Kaiqiang Song, Sangwoo Cho, Hongwei Wang et al.ACL 2023 · 4 citations
Related papers
- Filling the Gaps: Selective Knowledge Augmentation for LLM RecommendersJaehyun Lee, Sanghwan Jang, Seongku Kang, Hwanjo YuSIGIR 2026
- PepRec: Progressive Enhancement of Prompting for RecommendationYakun Yu, Shiang Qi, Baochun Li, Di NiuEMNLP 2024 · 2 citations
- Transparent and Scrutable Recommendations Using Natural Language User ProfilesJerome Ramos, Hossein A. Rahmani, Xi Wang, Xiao Fu et al.ACL 2024
- Refining Text Generation for Realistic Conversational Recommendation via Direct Preference OptimizationManato Tajiri, Michimasa InabaEMNLP 2025
- LWGR: Lagrangian-Constrained Personalized World Knowledge for Generative RecommendationLingyu Mu, Hao Deng, Haibo Xing, Kaican Lin et al.SIGIR 2026
