RaCT: Toward Amortized Ranking-Critical Training For Collaborative Filtering
Sam Lobel, Chunyuan Li, Jianfeng Gao, Lawrence Carin
Abstract
We investigate new methods for training collaborative filtering models based on actor-critic reinforcement learning, to more directly maximize ranking-based objective functions. Specifically, we train a critic network to approximate ranking-based metrics, and then update the actor network to directly optimize against the learned metrics. In contrast to traditional learning-to-rank methods that require re-running the optimization procedure for new lists, our critic-based method amortizes the scoring process with a neural network, and can directly provide the (approximate) ranking scores for new lists. We demonstrate the actor-critic's ability to significantly improve the performance of a variety of prediction models, and achieve better or comparable performance to a variety of strong baselines on three large-scale datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 83cfde06-2d2b-4de0-898e-9421ed5a727cCited by top-tier papers5
- Stochastic-Expert Variational Autoencoder for Collaborative FilteringYoon-Sik Cho, Min-hwan OhWWW 2022 · 15 citations
- It's Enough: Relaxing Diagonal Constraints in Linear Autoencoders for RecommendationJaewan Moon, Hye-young Kim, Jongwuk LeeSIGIR 2023 · 3 citations
- Leveraging User Behavior History for Personalized Email SearchKeping Bi, Pavel Metrikov, Chunyuan Li, Byungki ByunWWW 2021 · 3 citations
- Why is Normalization Necessary for Linear Recommenders?Seongmin Park, Mincheol Yoon, Hye-young Kim, Jongwuk LeeSIGIR 2025 · 1 citation
- ImplicitSLIM and How it Improves Embedding-based Collaborative FilteringIlya Shenbin, Sergey I. NikolenkoICLR 2024
Related papers
- Decision-Aware Actor-Critic with Function Approximation and Theoretical GuaranteesSharan Vaswani, Amirreza Kazemi, Reza Babanezhad Harikandeh, Nicolas Le RouxNeurIPS 2023 · 6 citations
- ACC-Collab: An Actor-Critic Approach to Multi-Agent LLM CollaborationAndrew Estornell, Jean-Francois Ton, Yuanshun Yao, Yang LiuICLR 2025 · 1 citation
- Towards Off-Policy Learning for Ranking Policies with Logged FeedbackTeng Xiao, Suhang WangAAAI 2022 · 8 citations
- Bringing Fairness to Actor-Critic Reinforcement Learning for Network Utility OptimizationJingdi Chen, Yimeng Wang, Tian LanINFOCOM 2021 · 23 citations
- An Efficient Combinatorial Optimization Model Using Learning-to-Rank DistillationHonguk Woo, Hyunsung Lee, Sangwoo ChoAAAI 2022 · 7 citations
