A Multi-Agent View of Wireless Video Streaming with Delayed Client-Feedback
Nouman Khan, Ujwal Dinesha, Subrahmanyam Arunachalam, Dheeraj Narasimha, Vijay G. Subramanian, Srinivas Shakkottai
摘要
We study the optimal control of multiple video streams over a wireless downlink from a base-transceiver-station (BTS)/access point to N end-devices (EDs). The BTS sends video packets to each ED under a joint transmission energy constraint, the EDs choose when to play out the received packets, and the collective goal is to provide a high Quality-of-Experience (QoE) to the clients/end-users. All EDs send feedback about their states and actions to the BTS which reaches it after a fixed deterministic delay. We analyze this team problem with delayed feedback as a cooperative Multi-Agent Constrained Partially Observable Markov Decision Process (MA-C-POMDP).
First, using a recently established strong duality result for MA-C-POMDPs, the original problem is decomposed into N independent unconstrained transmitter-receiver (two-agent) problemsall sharing a Lagrange multiplier (that also needs to be optimized for optimal control). Thereafter, the common information (CI) approach and the formalism of approximate information states (AISs) are used to guide the design of a neural-network based architecture for learning-based multi-agent control in a single unconstrained transmitter-receiver problem. Finally, simulations on a single transmitter-receiver pair with a stylized QoE model are performed to highlight the advantage of delay-aware two-agent coordination over the transmitter choosing both transmission and play-out actions (perceiving the delayed state of the receiver as its current state).
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
相关 Paper
- Multi-agent active perception with prediction rewardsMikko Lauri, Frans A. OliehoekNeurIPS 2020 · 被引用 13 次
- Semi-Online Precoding with Information Parsing for Cooperative MIMO Wireless NetworksJuncheng Wang, Ben Liang, Min Dong, Gary Boudreau 等INFOCOM 2022 · 被引用 1 次
- Queue-Learning: A Reinforcement Learning Approach for Providing Quality of ServiceMajid Raeis, Ali Tizghadam, Alberto Leon-GarciaAAAI 2021 · 被引用 25 次
- Unifying AoI Minimization and Remote Estimation - Optimal Sensor/Controller Coordination with Random Two-way DelayCho-Hsin Tsai, Chih-Chun WangINFOCOM 2020 · 被引用 29 次
- PIPHEN: Physical Interaction Prediction with Hamiltonian Energy NetworksKewei Chen, Yayu Long, Mingsheng ShangAAAI 2026
