Mowgli: Passively Learned Rate Control for Real-Time Video
Neil Agarwal, Rui Pan, Francis Y. Yan, Ravi Netravali
Abstract
Rate control algorithms are at the heart of video conferencing platforms, determining target bitrates that match dynamic network characteristics for high quality. Recent data-driven strategies have shown promise for this challenging task, but the performance degradation they introduce during training has been a nonstarter for many production services, precluding adoption. This paper aims to bolster the practicality of data-driven rate control by presenting an alternative avenue for experiential learning: leveraging purely existing telemetry logs produced by the incumbent algorithm in production. We observe that these logs contain effective decisions, although often at the wrong times or in the wrong order. To realize this approach despite the inherent uncertainty that log-based learning brings (i.e., lack of feedback for new decisions), our system, Mowgli, combines a variety of robust learning techniques (i.e., conservatively reasoning about alternate behavior to minimize risk and using a richer model formulation to account for environmental noise). Across diverse networks (emulated and real-world), Mowgli outperforms the widely deployed GCC algorithm, increasing average video bitrates by 15-39% while reducing freeze rates by 60-100%.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 015bfd9a-985e-4197-98ff-83c90a0daa36Cited by top-tier papers1
Ask how each one uses itBuilds on12
- Conservative Q-Learning for Offline Reinforcement LearningAviral Kumar, Aurick Zhou, George Tucker, Sergey LevineNeurIPS 2020 · 2,881 citations
- Critic Regularized RegressionZiyu Wang, Alexander Novikov, Konrad Zolna, Josh Merel et al.NeurIPS 2020 · 406 citations
- Learning in situ: a randomized experiment in video streamingFrancis Y. Yan, Hudson Ayers, Chenzhi Zhu, Sadjad Fouladi et al.NSDI 2020 · 360 citations
- Classic Meets Modern: a Pragmatic Learning-Based Congestion Control for the InternetSoheil Abbasloo, Chen-Yu Yen, H. Jonathan ChaoSIGCOMM 2020 · 257 citations
- OnRL: improving mobile video telephony via online reinforcement learningHuanhuan Zhang, Anfu Zhou, Jiamin Lu, Ruoxuan Ma et al.MobiCom 2020 · 105 citations
Related papers
- Computers Can Learn from the Heuristic Designs and Master Internet Congestion ControlChen-Yu Yen, Soheil Abbasloo, H. Jonathan ChaoSIGCOMM 2023 · 73 citations
- R-FEC: RL-based FEC Adjustment for Better QoE in WebRTCInsoo Lee, Seyeon Kim, Sandesh Dhawaskar Sathyanarayana, Kyungmin Bin et al.ACM MM 2022 · 37 citations
- An Intelligent Learning Approach to Achieve Near-Second Low-Latency Live Video Streaming under Highly Fluctuating NetworksGuanghui Zhang, Ke Liu, Mengbai Xiao, Bingshu Wang et al.ACM MM 2023 · 5 citations
- Learned Internet Congestion Control for Short Video UploadingTianchi Huang, Chao Zhou, Lianchen Jia, Rui-Xiao Zhang et al.ACM MM 2022 · 9 citations
- Mamba: Bringing Multi-Dimensional ABR to WebRTCYueheng Li, Zicheng Zhang, Hao Chen, Zhan MaACM MM 2023 · 15 citations
