OnRL: improving mobile video telephony via online reinforcement learning
Huanhuan Zhang, Anfu Zhou, Jiamin Lu, Ruoxuan Ma, Yuhan Hu, Cong Li, Xinyu Zhang, Huadong Ma, Xiaojiang Chen
Abstract
Machine learning models, particularly reinforcement learning (RL), have demonstrated great potential in optimizing video streaming applications. However, the state-of-the-art solutions are limited to an "offline learning" paradigm, i.e., the RL models are trained in simulators and then are operated in real networks. As a result, they inevitably suffer from the simulation-to-reality gap, showing far less satisfactory performance under real conditions compared with simulated environment. In this work, we close the gap by proposing OnRL, an online RL framework for real-time mobile video telephony. OnRL puts many individual RL agents directly into the video telephony system, which make video bitrate decisions in real-time and evolve their models over time. OnRL then aggregates these agents to form a high-level RL model that can help each individual to react to unseen network conditions. Moreover, OnRL incorporates novel mechanisms to handle the adverse impacts of inherent video traffic dynamics, and to eliminate risks of quality degradation caused by the RL model's exploration attempts. We implement OnRL on a mainstream operational video telephony system, Alibaba Taobao-live. In a month-long evaluation with 543 hours of video sessions from 151 real-world mobile users, OnRL outperforms the prior algorithms significantly, reducing video stalling rate by 14.22% while maintaining similar video quality.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get c697ff5b-8b82-4aa8-841d-99b05386f90cCited by top-tier papers18
- LiveNet: a low-latency video transport network for large-scale live streamingJinyang Li, Zhenyu Li, Ri Lu, Kai Xiao et al.SIGCOMM 2022 · 59 citations
- GRACE: Loss-Resilient Real-Time Video through Neural CodecsYihua Cheng, Ziyi Zhang, Hanchen Li, Anton Arapin et al.NSDI 2024 · 53 citations
- A Workload-Aware DVFS Robust to Concurrent Tasks for Mobile DevicesChengdong Lin, Kun Wang, Zhenjiang Li, Yu PuMobiCom 2023 · 52 citations
- Enabling High Quality Real-Time Communications with Adaptive Frame-RateZili Meng, Tingfeng Wang, Yixin Shen, Bo Wang et al.NSDI 2023 · 46 citations
- Pudica: Toward Near-Zero Queuing Delay in Congestion Control for Cloud GamingShibo Wang, Shusen Yang, Xiao Kong, Chenglei Wu et al.NSDI 2024 · 32 citations
Related papers
- From Ember to Blaze: Swift Interactive Video Adaptation via Meta-Reinforcement LearningXuedou Xiao, Mingxuan Yan, Yingying Zuo, Boxi Liu et al.INFOCOM 2023 · 14 citations
- Exstream: A Delay-minimized Streaming System with Explicit Frame Queueing Delay MeasurementShinik Park, Sanghyun Han, Junseon Kim, Jongyun Lee et al.INFOCOM 2024 · 1 citation
- AraLive: Automatic Reward Adaption for Learning-based Live Video StreamingHuanhuan Zhang, Liu zhuo, Haotian Li, Anfu Zhou et al.ACM MM 2024 · 6 citations
- EAVS: Edge-assisted Adaptive Video Streaming with Fine-grained Serverless PipelinesBiao Hou, Song Yang, Fernando A. Kuipers, Lei Jiao et al.INFOCOM 2023 · 26 citations
- Loki: improving long tail performance of learning-based real-time video adaptation by fusing rule-based modelsHuanhuan Zhang, Anfu Zhou, Yuhan Hu, Chaoyue Li et al.MobiCom 2021 · 78 citations
