Multi-Agent Distributed Reinforcement Learning for Making Decentralized Offloading Decisions
Jing Tan, Ramin Khalili, Holger Karl, Artur Hecker
摘要
We formulate computation offloading as a decentralized decision-making problem with autonomous agents. We design an interaction mechanism that incentivizes agents to align private and system goals by balancing between competition and cooperation. The mechanism provably has Nash equilibria with optimal resource allocation in the static case. For a dynamic environment, we propose a novel multi-agent online learning algorithm that learns with partial, delayed and noisy state information, and a reward signal that reduces information need to a great extent. Empirical results confirm that through learning, agents significantly improve both system and individual performance, e.g., 40% offloading failure rate reduction, 32% communication overhead reduction, up to 38% computation resource savings in low contention, 18% utilization increase with reduced load variation in high contention, and improvement in fairness. Results also confirm the algorithm’s good convergence and generalization property in significantly different environments.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper2
- Vehicular and Edge Computing for Emerging Connected and Autonomous Vehicle ApplicationsSabur Baidya, Yu-Jen Ku, Hengyu Zhao, Jishen Zhao 等DAC 2020 · 被引用 37 次
- Letting off STEAM: Distributed Runtime Traffic Scheduling for Service Function ChainingMarcel Blöcher, Ramin Khalili, Lin Wang, Patrick EugsterINFOCOM 2020 · 被引用 19 次
相关 Paper
- Decentralized Task Offloading in Edge Computing: A Multi-User Multi-Armed Bandit ApproachXiong Wang, Jiancheng Ye, John C. S. LuiINFOCOM 2022 · 被引用 89 次
- Socially-Optimal Mechanism Design for Incentivized Online LearningZhiyuan Wang, Lin Gao, Jianwei HuangINFOCOM 2022 · 被引用 11 次
- Online Learning for Load Balancing of Unknown Monotone Resource Allocation GamesIlai Bistritz, Nicholas BambosICML 2021 · 被引用 9 次
- Time Fairness in Online Knapsack ProblemsAdam Lechowicz, Rik Sengupta, Bo Sun, Shahin Kamali 等ICLR 2024 · 被引用 8 次
- ExplabOff: Towards Explorative and Collaborative Task Offloading via Mutual Information-Enhanced MARLTao Ren, Zheyuan Hu, Jianwei Niu, Yiming YaoINFOCOM 2025 · 被引用 2 次
