Mamba: Bringing Multi-Dimensional ABR to WebRTC
Yueheng Li, Zicheng Zhang, Hao Chen, Zhan Ma
Abstract
Contemporary real-time video communication systems, such as WebRTC, use an adaptive bitrate (ABR) algorithm to assure high-quality and low-delay services, e.g., promptly adjusting video bitrate according to the instantaneous network bandwidth. However, target bitrate decisions in the network and bitrate control in the codec are typically incoordinated and simply ignoring the effect of inappropriate resolution and frame rate settings also leads to compromised results in bitrate control, thus devastatingly deteriorating the quality of experience (QoE). To tackle these challenges, Mamba, an end-to-end multi-dimensional ABR algorithm is proposed, which utilizes multi-agent reinforcement learning (MARL) to maximize the user's QoE by adaptively and collaboratively adjusting encoding factors including the quantization parameters (QP), resolution, and frame rate based on observed states such as network conditions and video complexity information in a video conferencing system. We also introduce curriculum learning to improve the training efficiency of MARL. Both the in-lab and real-world evaluation results demonstrate the remarkable efficacy of Mamba.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 10daeb35-17c3-4669-bfe2-1acd4c42a954Cited by top-tier papers3
- Optimizing Adaptive Video Streaming with Human FeedbackTianchi Huang, Rui-Xiao Zhang, Chenglei Wu, Lifeng SunACM MM 2023 · 26 citations
- Harnessing WebRTC for Large-Scale Live StreamingWei Zhang, Tong Meng, Xianhua Zeng, Wei Yang et al.SIGCOMM 2025 · 4 citations
- PDStream: Slashing Long- Tail Delay in Interactive Video Streaming via Pseudo-Dual StreamingXuedou Xiao, Yingying Zuo, Mingxuan Yan, Kezhong Liu et al.INFOCOM 2025 · 1 citation
Builds on5
- A variegated look at 5G in the wild: performance, power, and QoE implicationsArvind Narayanan, Xumiao Zhang, Ruiyang Zhu, Ahmad Hassan et al.SIGCOMM 2021 · 259 citations
- Classic Meets Modern: a Pragmatic Learning-Based Congestion Control for the InternetSoheil Abbasloo, Chen-Yu Yen, H. Jonathan ChaoSIGCOMM 2020 · 257 citations
- From Few to More: Large-Scale Dynamic Multiagent Curriculum LearningWeixun Wang, Tianpei Yang, Yong Liu, Jianye Hao et al.AAAI 2020 · 138 citations
- CM3: Cooperative Multi-goal Multi-stage Multi-agent Reinforcement LearningJiachen Yang, Alireza Nakhaei, David Isele, Kikuo Fujimura et al.ICLR 2020 · 86 citations
- Loki: improving long tail performance of learning-based real-time video adaptation by fusing rule-based modelsHuanhuan Zhang, Anfu Zhou, Yuhan Hu, Chaoyue Li et al.MobiCom 2021 · 78 citations
Related papers
- R-FEC: RL-based FEC Adjustment for Better QoE in WebRTCInsoo Lee, Seyeon Kim, Sandesh Dhawaskar Sathyanarayana, Kyungmin Bin et al.ACM MM 2022 · 37 citations
- MultiLive: Adaptive Bitrate Control for Low-delay Multi-party Interactive Live StreamingZiyi Wang, Yong Cui, Xiaoyu Hu, Xin Wang et al.INFOCOM 2020 · 19 citations
- Meta Reinforcement Learning for Rate AdaptationAbdelhak Bentaleb, May Lim, Mehmet N. Akcay, Ali C. Begen et al.INFOCOM 2023 · 14 citations
- Buffer Awareness Neural Adaptive Video Streaming for Avoiding Extra Buffer ConsumptionTianchi Huang, Chao Zhou, Rui-Xiao Zhang, Chenglei Wu et al.INFOCOM 2023 · 27 citations
- AraLive: Automatic Reward Adaption for Learning-based Live Video StreamingHuanhuan Zhang, Liu zhuo, Haotian Li, Anfu Zhou et al.ACM MM 2024 · 6 citations
