Karma: Adaptive Video Streaming via Causal Sequence Modeling
Bowei Xu, Hao Chen, Zhan Ma
摘要
Optimal adaptive bitrate (ABR) decision depends on a comprehensive characterization of state transitions that involve interrelated modalities over time including environmental observations, returns, and actions. However, state-of-the-art learning-based ABR algorithms solely rely on past observations to decide the next action. This paradigm tends to cause a chain of deviations from optimal action when encountering unfamiliar observations, which consequently undermines the model generalization.
This paper presents Karma, an ABR algorithm that utilizes causal sequence modeling to improve generalization by comprehending the interrelated causality among past observations, returns, and actions and timely refining action when deviation occurs. Unlike direct observation-to-action mapping, Karma recurrently maintains a multi-dimensional time series of observations, returns, and actions as input and employs causal sequence modeling via a decision transformer to determine the next action. In the input sequence, Karma uses the maximum cumulative future quality of experience (QoE) (a.k.a, QoE-to-go) as an extended return signal, which is periodically estimated based on current network conditions and playback status. We evaluate Karma through trace-driven simulations and real-world field tests, demonstrating superior performance compared to existing state-of-the-art ABR algorithms, with an average QoE improvement ranging from 10.8% to 18.7% across diverse network conditions. Furthermore, Karma exhibits strong generalization capabilities, showing leading performance under unseen networks in both simulations and real-world tests.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper5
- Zero-Shot Text-to-Image GenerationAditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray 等ICML 2021 · 被引用 6,356 次
- Decision Transformer: Reinforcement Learning via Sequence ModelingLili Chen, Kevin Lu, Aravind Rajeswaran, Kimin Lee 等NeurIPS 2021 · 被引用 2,557 次
- Offline Reinforcement Learning as One Big Sequence Modeling ProblemMichael Janner, Qiyang Li, Sergey LevineNeurIPS 2021 · 被引用 950 次
- Genet: automatic curriculum generation for learning adaptation in networkingZhengxu Xia, Yajie Zhou, Francis Y. Yan, Junchen JiangSIGCOMM 2022 · 被引用 57 次
- Improving Generalization for Neural Adaptive Video Streaming via Meta Reinforcement LearningNuowen Kan, Yuankun Jiang, Chenglin Li, Wenrui Dai 等ACM MM 2022 · 被引用 48 次
相关 Paper
- Progressive Learning with Human Feedback for Personalized Adaptive Video StreamingZhaohui Jiang, Xuening Feng, Tianchi Huang, Ruixiao Zhang 等ACM MM 2025
- Optimizing Adaptive Video Streaming with Human FeedbackTianchi Huang, Rui-Xiao Zhang, Chenglei Wu, Lifeng SunACM MM 2023 · 被引用 26 次
- AraLive: Automatic Reward Adaption for Learning-based Live Video StreamingHuanhuan Zhang, Liu zhuo, Haotian Li, Anfu Zhou 等ACM MM 2024 · 被引用 6 次
- Buffer Awareness Neural Adaptive Video Streaming for Avoiding Extra Buffer ConsumptionTianchi Huang, Chao Zhou, Rui-Xiao Zhang, Chenglei Wu 等INFOCOM 2023 · 被引用 27 次
- SODA: An Adaptive Bitrate Controller for Consistent High-Quality Video StreamingTianyu Chen, Yiheng Lin, Nicolas Christianson, Zahaib Akhtar 等SIGCOMM 2024 · 被引用 33 次
