Stochastic Primal-Dual Decoding for Multiobjective Generative Recommender Systems
Dmitrii Moor, Ben Carterette, Senthilkumar Krishnamoorthy, Kyle Kretschman, Denis Beslic, Melissa Yalla, Alice Y. Wang, Mounia Lalmas
Abstract
Recent advances in recommender systems (RS) have shown substantial performance gains through generative modelling. In practice, recommendation often involves constructing slates--ordered lists of items--that must satisfy multiple objectives beyond relevance, such as constraints defined over item attributes or fairness constraints. Existing multiobjective approaches either rely on post-processing techniques designed for non-generative settings, or incorporate auxiliary objectives directly into model training. The former does not explicitly account for the sequential nature of generative RS, while the latter is often impractical in large-scale systems. We propose a lightweight, inference-time decoding layer that augments autoregressive generative RS to support multiobjective slate generation without modifying or retraining the underlying model. We formulate decoding as an online constrained optimisation problem, where items are selected sequentially, and trade-offs between relevance and auxiliary objectives are adjusted dynamically based on the remaining constraint slack, i.e., how much of each objective remains to be satisfied. This is implemented via a stochastic primal-dual approximation scheme that balances relevance and auxiliary objectives during generation. We provide theoretical guarantees on constraint violation and regret, and evaluate the proposed approach through extensive offline experiments and a large-scale online A/B experiment in a real-world recommender system. Our results show consistent improvements in multiobjective trade-offs, including a +1.8% gain in the auxiliary objectives achieved at zero cost to user satisfaction.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 983cde46-b9d8-4034-8539-a798dbb9f051Builds on10
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Decision Transformer: Reinforcement Learning via Sequence ModelingLili Chen, Kevin Lu, Aravind Rajeswaran, Kimin Lee et al.NeurIPS 2021 · 2,557 citations
- Simple and Effective Masked Diffusion Language ModelsSubham S. Sahoo, Marianne Arriola, Yair Schiff, Aaron Gokaslan et al.NeurIPS 2024 · 929 citations
- Algorithmic Effects on the Diversity of Consumption on SpotifyAshton Anderson, Lucas Maystre, Ian Anderson, Rishabh Mehrotra et al.WWW 2020 · 211 citations
- Actions Speak Louder than Words: Trillion-Parameter Sequential Transducers for Generative RecommendationsJiaqi Zhai, Lucy Liao, Xing Liu, Yueming Wang et al.ICML 2024 · 200 citations
Related papers
- APAO: Bridging the Training-Inference Gap in Generative Recommendation via Adaptive Prefix-Aware OptimizationYuanqing Yu, Yifan Wang, Weizhi Ma, Zhiqiang Guo et al.KDD 2026 · 4 citations
- User-item fairness tradeoffs in recommendationsSophie Greenwood, Sudalakshmee Chiniah, Nikhil GargNeurIPS 2024 · 15 citations
- The NodeHopper: Enabling Low Latency Ranking with Constraints via a Fast Dual SolverAnton Zhernov, Krishnamurthy (Dj) Dvijotham, Ivan Lobov, Dan A. Calian et al.KDD 2020 · 2 citations
- Interpolating Item and User Fairness in Multi-Sided RecommendationsQinyi Chen, Jason Cheuk Nam Liang, Negin Golrezaei, Djallel BouneffoufNeurIPS 2024 · 8 citations
- Multi-slots Online Matching with High EntropyXingyu Lu, Qintong Wu, Wenliang ZhongICML 2022 · 3 citations
