RaCT: Toward Amortized Ranking-Critical Training For Collaborative Filtering
Sam Lobel, Chunyuan Li, Jianfeng Gao, Lawrence Carin
摘要
We investigate new methods for training collaborative filtering models based on actor-critic reinforcement learning, to more directly maximize ranking-based objective functions. Specifically, we train a critic network to approximate ranking-based metrics, and then update the actor network to directly optimize against the learned metrics. In contrast to traditional learning-to-rank methods that require re-running the optimization procedure for new lists, our critic-based method amortizes the scoring process with a neural network, and can directly provide the (approximate) ranking scores for new lists. We demonstrate the actor-critic's ability to significantly improve the performance of a variety of prediction models, and achieve better or comparable performance to a variety of strong baselines on three large-scale datasets.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Stochastic-Expert Variational Autoencoder for Collaborative FilteringYoon-Sik Cho, Min-hwan OhWWW 2022 · 被引用 15 次
- It's Enough: Relaxing Diagonal Constraints in Linear Autoencoders for RecommendationJaewan Moon, Hye-young Kim, Jongwuk LeeSIGIR 2023 · 被引用 3 次
- Leveraging User Behavior History for Personalized Email SearchKeping Bi, Pavel Metrikov, Chunyuan Li, Byungki ByunWWW 2021 · 被引用 3 次
- Why is Normalization Necessary for Linear Recommenders?Seongmin Park, Mincheol Yoon, Hye-young Kim, Jongwuk LeeSIGIR 2025 · 被引用 1 次
- ImplicitSLIM and How it Improves Embedding-based Collaborative FilteringIlya Shenbin, Sergey I. NikolenkoICLR 2024
相关 Paper
- Decision-Aware Actor-Critic with Function Approximation and Theoretical GuaranteesSharan Vaswani, Amirreza Kazemi, Reza Babanezhad Harikandeh, Nicolas Le RouxNeurIPS 2023 · 被引用 6 次
- ACC-Collab: An Actor-Critic Approach to Multi-Agent LLM CollaborationAndrew Estornell, Jean-Francois Ton, Yuanshun Yao, Yang LiuICLR 2025 · 被引用 1 次
- Towards Off-Policy Learning for Ranking Policies with Logged FeedbackTeng Xiao, Suhang WangAAAI 2022 · 被引用 8 次
- Bringing Fairness to Actor-Critic Reinforcement Learning for Network Utility OptimizationJingdi Chen, Yimeng Wang, Tian LanINFOCOM 2021 · 被引用 23 次
- An Efficient Combinatorial Optimization Model Using Learning-to-Rank DistillationHonguk Woo, Hyunsung Lee, Sangwoo ChoAAAI 2022 · 被引用 7 次
