Mamba: Bringing Multi-Dimensional ABR to WebRTC
Yueheng Li, Zicheng Zhang, Hao Chen, Zhan Ma
摘要
Contemporary real-time video communication systems, such as WebRTC, use an adaptive bitrate (ABR) algorithm to assure high-quality and low-delay services, e.g., promptly adjusting video bitrate according to the instantaneous network bandwidth. However, target bitrate decisions in the network and bitrate control in the codec are typically incoordinated and simply ignoring the effect of inappropriate resolution and frame rate settings also leads to compromised results in bitrate control, thus devastatingly deteriorating the quality of experience (QoE). To tackle these challenges, Mamba, an end-to-end multi-dimensional ABR algorithm is proposed, which utilizes multi-agent reinforcement learning (MARL) to maximize the user's QoE by adaptively and collaboratively adjusting encoding factors including the quantization parameters (QP), resolution, and frame rate based on observed states such as network conditions and video complexity information in a video conferencing system. We also introduce curriculum learning to improve the training efficiency of MARL. Both the in-lab and real-world evaluation results demonstrate the remarkable efficacy of Mamba.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Optimizing Adaptive Video Streaming with Human FeedbackTianchi Huang, Rui-Xiao Zhang, Chenglei Wu, Lifeng SunACM MM 2023 · 被引用 26 次
- Harnessing WebRTC for Large-Scale Live StreamingWei Zhang, Tong Meng, Xianhua Zeng, Wei Yang 等SIGCOMM 2025 · 被引用 4 次
- PDStream: Slashing Long- Tail Delay in Interactive Video Streaming via Pseudo-Dual StreamingXuedou Xiao, Yingying Zuo, Mingxuan Yan, Kezhong Liu 等INFOCOM 2025 · 被引用 1 次
它引用的顶会 Paper5
- A variegated look at 5G in the wild: performance, power, and QoE implicationsArvind Narayanan, Xumiao Zhang, Ruiyang Zhu, Ahmad Hassan 等SIGCOMM 2021 · 被引用 259 次
- Classic Meets Modern: a Pragmatic Learning-Based Congestion Control for the InternetSoheil Abbasloo, Chen-Yu Yen, H. Jonathan ChaoSIGCOMM 2020 · 被引用 257 次
- From Few to More: Large-Scale Dynamic Multiagent Curriculum LearningWeixun Wang, Tianpei Yang, Yong Liu, Jianye Hao 等AAAI 2020 · 被引用 138 次
- CM3: Cooperative Multi-goal Multi-stage Multi-agent Reinforcement LearningJiachen Yang, Alireza Nakhaei, David Isele, Kikuo Fujimura 等ICLR 2020 · 被引用 86 次
- Loki: improving long tail performance of learning-based real-time video adaptation by fusing rule-based modelsHuanhuan Zhang, Anfu Zhou, Yuhan Hu, Chaoyue Li 等MobiCom 2021 · 被引用 78 次
相关 Paper
- R-FEC: RL-based FEC Adjustment for Better QoE in WebRTCInsoo Lee, Seyeon Kim, Sandesh Dhawaskar Sathyanarayana, Kyungmin Bin 等ACM MM 2022 · 被引用 37 次
- MultiLive: Adaptive Bitrate Control for Low-delay Multi-party Interactive Live StreamingZiyi Wang, Yong Cui, Xiaoyu Hu, Xin Wang 等INFOCOM 2020 · 被引用 19 次
- Meta Reinforcement Learning for Rate AdaptationAbdelhak Bentaleb, May Lim, Mehmet N. Akcay, Ali C. Begen 等INFOCOM 2023 · 被引用 14 次
- Buffer Awareness Neural Adaptive Video Streaming for Avoiding Extra Buffer ConsumptionTianchi Huang, Chao Zhou, Rui-Xiao Zhang, Chenglei Wu 等INFOCOM 2023 · 被引用 27 次
- AraLive: Automatic Reward Adaption for Learning-based Live Video StreamingHuanhuan Zhang, Liu zhuo, Haotian Li, Anfu Zhou 等ACM MM 2024 · 被引用 6 次
