From Ember to Blaze: Swift Interactive Video Adaptation via Meta-Reinforcement Learning
Xuedou Xiao, Mingxuan Yan, Yingying Zuo, Boxi Liu, Paul Ruan, Yang Cao, Wei Wang
摘要
Maximizing quality of experience (QoE) for interactive video streaming has been a long-standing challenge, as its delay-sensitive nature makes it more vulnerable to bandwidth fluctuations. While reinforcement learning (RL) has demonstrated great potential, existing works are either limited by fixed models or require enormous data/time for online adaptation, which struggle to fit time-varying and diverse network states. Driven by these practical concerns, we perform large-scale measurements on WeChat for Business’s interactive video service to study real-world network fluctuations. Surprisingly, our analysis shows that, compared to time-varying network metrics, network sequences exhibit noticeable short-term continuity, sufficient for few-shot learning requirement. We thus propose Fiammetta, the first meta-RL-based bitrate adaptation algorithm for interactive video streaming. Building on the short-term continuity, Fiammetta accumulates learning experiences through offline meta-training and enables fast online adaptation to changing network states through few gradient updates. Moreover, Fiammetta innovatively incorporates a probing mechanism for real-time monitoring of network states, and proposes an adaptive meta-testing mechanism for seamless adaptation. We implement Fiammetta on a testbed whose end-to-end network follows the real-world WeChat for Business traces. The results show that Fiammetta outperforms prior algorithms significantly, improving video bitrate by 3.6%-16.2% without increasing stalling rate.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper6
- Learning in situ: a randomized experiment in video streamingFrancis Y. Yan, Hudson Ayers, Chenzhi Zhu, Sadjad Fouladi 等NSDI 2020 · 被引用 360 次
- Meta-Learning with Task-Adaptive Loss Function for Few-Shot LearningSungyong Baik, Janghoon Choi, Heewon Kim, Dohee Cho 等ICCV 2021 · 被引用 146 次
- OnRL: improving mobile video telephony via online reinforcement learningHuanhuan Zhang, Anfu Zhou, Jiamin Lu, Ruoxuan Ma 等MobiCom 2020 · 被引用 105 次
- Loki: improving long tail performance of learning-based real-time video adaptation by fusing rule-based modelsHuanhuan Zhang, Anfu Zhou, Yuhan Hu, Chaoyue Li 等MobiCom 2021 · 被引用 78 次
- Adaptive Bitrate with User-level QoE Preference for Video StreamingXutong Zuo, Jiayu Yang, Mowei Wang, Yong CuiINFOCOM 2022 · 被引用 65 次
相关 Paper
- Improving Generalization for Neural Adaptive Video Streaming via Meta Reinforcement LearningNuowen Kan, Yuankun Jiang, Chenglin Li, Wenrui Dai 等ACM MM 2022 · 被引用 48 次
- Personalized 360-Degree Video Streaming: A Meta-Learning ApproachYiyun Lu, Yifei Zhu, Zhi WangACM MM 2022 · 被引用 28 次
- Meta Reinforcement Learning for Rate AdaptationAbdelhak Bentaleb, May Lim, Mehmet N. Akcay, Ali C. Begen 等INFOCOM 2023 · 被引用 14 次
- AraLive: Automatic Reward Adaption for Learning-based Live Video StreamingHuanhuan Zhang, Liu zhuo, Haotian Li, Anfu Zhou 等ACM MM 2024 · 被引用 6 次
- An Intelligent Learning Approach to Achieve Near-Second Low-Latency Live Video Streaming under Highly Fluctuating NetworksGuanghui Zhang, Ke Liu, Mengbai Xiao, Bingshu Wang 等ACM MM 2023 · 被引用 5 次
