Generative Flow Network for Listwise Recommendation
Shuchang Liu, Qingpeng Cai, Zhankui He, Bowen Sun, Julian J. McAuley, Dong Zheng, Peng Jiang, Kun Gai
Abstract
Personalized recommender systems fulfill the daily demands of customers and boost online businesses. The goal is to learn a policy that can generate a list of items that matches the user's demand or interest. While most existing methods learn a pointwise scoring model that predicts the ranking score of each individual item, recent research shows that the listwise approach can further improve the recommendation quality by modeling the intra-list correlations of items that are exposed together. This has motivated the recent list reranking and generative recommendation approaches that optimize the overall utility of the entire list. However, it is challenging to explore the combinatorial space of list actions and existing methods that use cross-entropy loss may suffer from low diversity issues. In this work, we aim to learn a policy that can generate sufficiently diverse item lists for users while maintaining high recommendation quality. The proposed solution, GFN4Rec, is a generative method that takes the insight of the flow network to ensure the alignment between list generation probability and its reward. The key advantages of our solution are the log scale reward matching loss that intrinsically improves the generation diversity and the autoregressive item selection model that captures the item mutual influences while capturing future reward of the list. As validation of our method's effectiveness and its superior diversity during active exploration, we conduct experiments on simulated online environments as well as an offline evaluation framework for two real-world datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 7729cd35-18b8-49a2-b324-e77cece12112Cited by top-tier papers14
- DIET: Customized Slimming for Incompatible Networks in Sequential RecommendationKairui Fu, Shengyu Zhang, Zheqi Lv, Jingyuan Chen et al.KDD 2024 · 6 citations
- AURO: Reinforcement Learning for Adaptive User Retention Optimization in Recommender SystemsZhenghai Xue, Qingpeng Cai, Bin Yang, Lantao Hu et al.WWW 2025 · 6 citations
- Comprehensive List Generation for Multi-Generator RerankingHailan Yang, Zhenyu Qi, Shuchang Liu, Xiaoyu Yang et al.SIGIR 2025 · 4 citations
- GoalRank: Group-Relative Optimization for a Large Ranking ModelKaike Zhang, Xiaobei Wang, Shuchang Liu, HailanYang et al.ICLR 2026 · 2 citations
- GFlowGR: Fine-tuning Generative Recommendation Frameworks with Generative Flow NetworksYejing Wang, Shengyu Zhou, Jinyu Lu, Qidong Liu et al.SIGIR 2026 · 1 citation
Builds on7
- Flow Network based Generative Models for Non-Iterative Diverse Candidate GenerationEmmanuel Bengio, Moksh Jain, Maksym Korablyov, Doina Precup et al.NeurIPS 2021 · 565 citations
- Trajectory balance: Improved credit assignment in GFlowNetsNikolay Malkin, Moksh Jain, Emmanuel Bengio, Chen Sun et al.NeurIPS 2022 · 316 citations
- Better Training of GFlowNets with Local Credit and Incomplete TrajectoriesLing Pan, Nikolay Malkin, Dinghuai Zhang, Yoshua BengioICML 2023 · 100 citations
- Two-Stage Constrained Actor-Critic for Short Video RecommendationQingpeng Cai, Zhenghai Xue, Chi Zhang, Wanqi Xue et al.WWW 2023 · 60 citations
- Exploration and Regularization of the Latent Action Space in RecommendationShuchang Liu, Qingpeng Cai, Bowen Sun, Yuhao Wang et al.WWW 2023 · 54 citations
Related papers
- QGFN: Controllable Greediness with Action ValuesElaine Lau, Stephen Zhewen Lu, Ling Pan, Doina Precup et al.NeurIPS 2024 · 21 citations
- Advantage-Conditioned Flow Policy for Offline Reinforcement Learning in RecommendationXiaocong Chen, Siyu Wang, Lina YaoSIGIR 2026
- Large-Scale Online Learning for Generative List Recommendation in E-commerce: An Environment Policy Optimization ApproachYuan Wang, Zhiyu Li, Ang Gao, Changshuo Zhang et al.SIGIR 2026
- Pessimistic Backward Policy for GFlowNetsHyosoon Jang, Yunhui Jang, Minsu Kim, Jinkyoo Park et al.NeurIPS 2024 · 14 citations
- Generative Augmented Flow NetworksLing Pan, Dinghuai Zhang, Aaron C. Courville, Longbo Huang et al.ICLR 2023 · 2 citations
