FlowTime: Towards Continuous Generative Watch Time Prediction via Flow-based Personalized Priors
Hongxu Ma, Han Zhou, Chenghou Jin, Jie Zhang, Xiaoyu Yang, Chunjie Chen, Jihong Guan, Shuigeng Zhou
Abstract
Watch time has emerged as a pivotal metric for optimizing deep user engagement in short-video recommender systems. However, current methods of watch time prediction (WTP) suffer from inherent paradigm-specific limitations. Direct Regression faces mean-collapse due to unimodal Gaussian assumptions, while Ordinal Regression is hampered by quantization errors from rigid discretization. Similarly, Discrete Generative Regression struggles with high inference latency and heuristic vocabulary design. Beyond these specific flaws, a shared deficiency is the inability to capture the intrinsic multimodality and heterogeneity of User-Item Interaction Patterns. To address these challenges, we first revisit the WTP problem from a causal perspective, and identify these user-specific patterns as structural confounders that modulate watch time outcomes, where identical interests manifest as distinct watch time outcomes conditioned on diverse user habits. Then, we formally propose a new (or the fourth) paradigm --- Continuous Generative Regression, and introduce FlowTime, a novel method utilizing a One-step Generative Variational Autoencoder. FlowTime effectively circumvents the latency of iterative denoising while maintaining the expressivity of continuous latent spaces. Furthermore, we design a Flow-based Personalized Prior that leverages NFs to warp a standard Gaussian prior into a complex, history-conditioned manifold, thereby enabling the adaptive modeling of multimodal interaction patterns. Finally, we build TimeRec, the first open-source WTP Library, alongside a novel personalization metric to establish a rigorous benchmarking standard. Extensive offline experiments and online A/B tests demonstrate FlowTime's significant superiority over SOTA methods. Our code is available at https://github.com/snailma0229/TimeRec.git.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f12175ed-424c-4d6c-afe8-842b2b91b272Builds on13
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Generate What You Prefer: Reshaping Sequential Recommendation via Guided DiffusionZhengyi Yang, Jiancan Wu, Zhicai Wang, Xiang Wang et al.NeurIPS 2023 · 205 citations
- Flow Matching for Generative ModelingYaron Lipman, Ricky T. Q. Chen, Heli Ben-Hamu, Maximilian Nickel et al.ICLR 2023 · 87 citations
- MS-DETR: Towards Effective Video Moment Retrieval and Highlight Detection by Joint Motion-Semantic LearningHongxu Ma, Guanshuo Wang, Fufu Yu, Qiong Jia et al.ACM MM 2025 · 9 citations
- Counteracting Duration Bias in Video Recommendation via Counterfactual Watch TimeHaiyuan Zhao, Guohao Cai, Jieming Zhu, Zhenhua Dong et al.KDD 2024 · 9 citations
Related papers
- Generative Regression Based Watch Time Prediction for Short-Video RecommendationHongxu Ma, Kai Tian, Tao Zhang, Xuefeng Zhang et al.WWW 2026 · 6 citations
- Calibrating Video Watch-time Predictions with Credible Prototype AlignmentChao Cui, Shisong Tang, Fan Li, Jiechao Gao et al.ICML 2025
- CREAD: A Classification-Restoration Framework with Error Adaptive Discretization for Watch Time Prediction in Video Recommender SystemsJie Sun, Zhaoying Ding, Xiaoshuang Chen, Qi Chen et al.AAAI 2024
- DVR: Micro-Video Recommendation Optimizing Watch-Time-Gain under Duration BiasYu Zheng, Chen Gao, Jingtao Ding, Lingling Yi et al.ACM MM 2022 · 28 citations
- An Action-Aware Generative Sequence Modeling for Short Video RecommendationWenhao Li, Zihan Lin, Zhengxiao Guo, Jie Zhou et al.SIGIR 2026
