Reinforcement Learning-based Congestion Control: A Systematic Evaluation of Fairness, Efficiency and Responsiveness
Luca Giacomoni, George Parisis
摘要
Reinforcement learning (RL)-based congestion control (CC) promises efficient CC in a fast-changing networking landscape, where evolving communication technologies, applications and traffic workloads pose severe challenges to human-derived, static CC algorithms. RL-based CC is in its early days and substantial research is required to understand existing limitations, identify research challenges and, eventually, yield deployable solutions for real-world networks. In this paper we present the first reproducible and systematic study of RL-based CC with the aim to highlight strengths and uncover fundamental limitations of the state-of-the-art. We identify challenges in evaluating RL-based CC, establish a methodology for studying said approaches and perform large-scale experimentation with RL-based CC approaches that are publicly available. We show that existing approaches can acquire all available bandwidth swiftly and are resistant to non-congestive loss, however, this is commonly at the cost of excessive packet loss in normal operation. We show that, as fairness is not embedded directly into reward functions, existing approaches exhibit unfairness in almost all tested network setups. Finally, we provide evidence that existing RL-based CC approaches under-perform when the available bandwidth and end-to-end latency dynamically change. Our experimentation codebase and datasets are publicly available with the aim to galvanise the community towards transparency and reproducibility, which have been recognised as crucial for researching and evaluating machine-generated policies.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper1
问问它们各自怎么用它相关 Paper
- Achieving Fairness Generalizability for Learning-based Congestion Control with JuryHan Tian, Xudong Liao, Decang Sun, Chaoliang Zeng 等EuroSys 2025 · 被引用 10 次
- Mutant: Learning Congestion Control from Existing Protocols via Online Reinforcement LearningLorenzo Pappone, Alessio Sacco, Flavio EspositoNSDI 2025 · 被引用 24 次
- DACC: Data Augmentation for Learning-based Congestion ControlXiaojun Zhu, Jiawei Huang, Haifeng Liu, Zhaoyi Li 等INFOCOM 2025 · 被引用 1 次
- Owl: Congestion Control with Partially Invisible Networks via Reinforcement LearningAlessio Sacco, Matteo Flocco, Flavio Esposito, Guido MarchettoINFOCOM 2021 · 被引用 40 次
- Marten: A Built-in Security DRL-Based Congestion Control Framework by Polishing the ExpertZhiyuan Pan, Jianer Zhou, Xinyi Qiu, Weichao Li 等INFOCOM 2023 · 被引用 8 次
