An Action-Aware Generative Sequence Modeling for Short Video Recommendation
Wenhao Li, Zihan Lin, Zhengxiao Guo, Jie Zhou, Shukai Liu, Yongqi Liu, Chuan Luo
Abstract
With the rapid development of the Internet, users have increasingly higher expectations for the recommendation accuracy of online content consumption platforms (e.g., short video platforms). However, short videos often contain diverse segments, and users may not hold the same attitude toward all of them (e.g., music enthusiasts may not enjoy all songs in a medley). Traditional binary-classification recommendation models, which treat a video as a single holistic entity, face limitations in accurately capturing such nuanced preferences. Considering that user consumption is a temporal process, this paper demonstrates that the timing of user actions can represent diverse intentions through statistical analysis and examination of action patterns. Based on this insight, we propose a novel modeling paradigm: Action-Aware Gen erative Sequence Network (A2Gen ), which refines user actions (e.g., Like and Follow, etc.) along the temporal dimension and chains them into sequences for unified processing and prediction. First, we introduce the Context-aware Attention Module (CAM) to model action sequences enriched with item-specific contextual features. Building upon this, we develop the Hierarchical Sequence Encoder (HSE) to learn temporal action patterns from users' historical actions. Finally, through leveraging CAM, we design a module for action sequence generation: the Action-seq Autoregressive Generator (AAG). Extensive offline experiments on the Kuaishou's dataset and the Tmall public dataset demonstrate the superiority of our proposed model. Furthermore, through large-scale online A/B testing deployed on Kuaishou's platform, our model achieves significant improvements over baseline methods in multi-task prediction by leveraging sequential information. Specifically, it yields increases of 0.34% in user watch time, 8.1% in interaction rate, and 0.162% in overall user retention (LifeTime-7), leading to successful deployment across all traffic, serving over 400 million users every day.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c47fb9b8-8efd-4e0b-a391-0e45428c70faBuilds on8
- Actions Speak Louder than Words: Trillion-Parameter Sequential Transducers for Generative RecommendationsJiaqi Zhai, Lucy Liao, Xing Liu, Yueming Wang et al.ICML 2024 · 200 citations
- Collaborative Large Language Model for Recommender SystemsYaochen Zhu, Liang Wu, Qi Guo, Liangjie Hong et al.WWW 2024 · 150 citations
- Adapting Large Language Models by Integrating Collaborative Semantics for RecommendationBowen Zheng, Yupeng Hou, Hongyu Lu, Yu Chen et al.ICDE 2024 · 132 citations
- LLaRA: Large Language-Recommendation AssistantJiayi Liao, Sihang Li, Zhengyi Yang, Jiancan Wu et al.SIGIR 2024 · 120 citations
- AdaTask: A Task-Aware Adaptive Learning Rate Approach to Multi-Task LearningEnneng Yang, Junwei Pan, Ximei Wang, Haibin Yu et al.AAAI 2023 · 70 citations
Related papers
- Generative Regression Based Watch Time Prediction for Short-Video RecommendationHongxu Ma, Kai Tian, Tao Zhang, Xuefeng Zhang et al.WWW 2026 · 6 citations
- Short Video Segment-level User Dynamic Interests Modeling in Personalized RecommendationZhiyu He, Zhixin Ling, Jiayu Li, Zhiqiang Guo et al.SIGIR 2025 · 4 citations
- What Aspect Do You Like: Multi-scale Time-aware User Interest Modeling for Micro-video RecommendationHao Jiang, Wenjie Wang, Yinwei Wei, Zan Gao et al.ACM MM 2020 · 65 citations
- Disentangling Long and Short-Term Interests for RecommendationYu Zheng, Chen Gao, Jianxin Chang, Yanan Niu et al.WWW 2022 · 128 citations
- Temporal-Series-Aware Adaptive Positional Encoding for Transformer-based Sequential RecommendationRongbo Qi, Yaqi Zhang, Chunyao Song, Tingjian GeWWW 2026
