Tunable LLM-based Proactive Recommendation Agent
Mingze Wang, Chongming Gao, Wenjie Wang, Yangyang Li, Fuli Feng
摘要
Recommender systems are indispensable on various digital platforms. However, traditional methods often reinforce existing user interests, which leads to echo chambers and limits diversity. Proactive Recommendation Systems (PRS) aim to address this issue by cultivating users’ latent interests through multi-step recommendations. Despite advancements, challenges persist particularly in optimizing long-term rewards and adapting to real-time user feed-back. In this study, we propose an LLM-based Actor-Critic Agent framework to enhance PRS. This framework utilizes the LLM-based agent to adjust recommendations in real time based on feedback and employs agent-tuning meth-ods to optimize long-term rewards using three proposed reward functions. Extensive experiments validate the significant superiority of this framework over existing methods by optimizing long-term rewards and dynamically evolving with user feedback. Our codes are available at https://github.com/gnaWeinrE/T-PRA .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Uncertainty-aware Generative RecommendationChenxiao Fan, Chongming Gao, Yaxin Gong, Haoyan Liu 等KDD 2026 · 被引用 2 次
- ProRL: Effective Reinforcement Learning for Proactive Recommendation via Rectified Policy Gradient EstimationHongru Hou, Tiehua Mei, Denghui Geng, Jinhui Huang 等ICML 2026
它引用的顶会 Paper19
- Retrieval-Augmented Generation for Knowledge-Intensive NLP TasksPatrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni 等NeurIPS 2020 · 被引用 19,162 次
- Direct Preference Optimization: Your Language Model is Secretly a Reward ModelRafael Rafailov, Archit Sharma, Eric Mitchell, Christopher D. Manning 等NeurIPS 2023 · 被引用 10,924 次
- Large Language Models are Zero-Shot ReasonersTakeshi Kojima, Shixiang Shane Gu, Machel Reid, Yutaka Matsuo 等NeurIPS 2022 · 被引用 8,168 次
- Reflexion: language agents with verbal reinforcement learningNoah Shinn, Federico Cassano, Ashwin Gopinath, Karthik Narasimhan 等NeurIPS 2023 · 被引用 5,828 次
- Tree of Thoughts: Deliberate Problem Solving with Large Language ModelsShunyu Yao, Dian Yu, Jeffrey Zhao, Izhak Shafran 等NeurIPS 2023 · 被引用 5,068 次
相关 Paper
- Adaptive Preference Arithmetic: A Personalized Agent with Adaptive Preference Arithmetic for Dynamic Preference ModelingHongyi Nie, Yaqing Wang, Mingyang Zhou, Feiyang Pan 等NeurIPS 2025 · 被引用 1 次
- ITMPRec: Intention-based Targeted Multi-round Proactive RecommendationYahong Lian, Chunyao Song, Tingjian GeWWW 2025 · 被引用 6 次
- RecCocktail: A Generalizable and Efficient Framework for LLM-Based RecommendationMin Hou, Chenxi Bai, Le Wu, Hao Liu 等AAAI 2026 · 被引用 2 次
- ProMax: Exploring the Potential of LLM-derived Profiles with Distribution Shaping for Recommender SystemsYi Zhang, Yiwen Zhang, Kai Zheng, Tong Chen 等SIGIR 2026
- PolicySim: An LLM-Based Agent Social Simulation Sandbox for Proactive Policy OptimizationRenhong Huang, Ning Tang, Jiarong Xu, Yuxuan Cao 等WWW 2026
