A Multi-Agent View of Wireless Video Streaming with Delayed Client-Feedback
Nouman Khan, Ujwal Dinesha, Subrahmanyam Arunachalam, Dheeraj Narasimha, Vijay G. Subramanian, Srinivas Shakkottai
Abstract
We study the optimal control of multiple video streams over a wireless downlink from a base-transceiver-station (BTS)/access point to N end-devices (EDs). The BTS sends video packets to each ED under a joint transmission energy constraint, the EDs choose when to play out the received packets, and the collective goal is to provide a high Quality-of-Experience (QoE) to the clients/end-users. All EDs send feedback about their states and actions to the BTS which reaches it after a fixed deterministic delay. We analyze this team problem with delayed feedback as a cooperative Multi-Agent Constrained Partially Observable Markov Decision Process (MA-C-POMDP).
First, using a recently established strong duality result for MA-C-POMDPs, the original problem is decomposed into N independent unconstrained transmitter-receiver (two-agent) problemsall sharing a Lagrange multiplier (that also needs to be optimized for optimal control). Thereafter, the common information (CI) approach and the formalism of approximate information states (AISs) are used to guide the design of a neural-network based architecture for learning-based multi-agent control in a single unconstrained transmitter-receiver problem. Finally, simulations on a single transmitter-receiver pair with a stylized QoE model are performed to highlight the advantage of delay-aware two-agent coordination over the transmitter choosing both transmission and play-out actions (perceiving the delayed state of the receiver as its current state).
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d2df4bda-f31f-40ee-ab6c-c271bbdd10a8Related papers
- Multi-agent active perception with prediction rewardsMikko Lauri, Frans A. OliehoekNeurIPS 2020 · 13 citations
- Semi-Online Precoding with Information Parsing for Cooperative MIMO Wireless NetworksJuncheng Wang, Ben Liang, Min Dong, Gary Boudreau et al.INFOCOM 2022 · 1 citation
- Queue-Learning: A Reinforcement Learning Approach for Providing Quality of ServiceMajid Raeis, Ali Tizghadam, Alberto Leon-GarciaAAAI 2021 · 25 citations
- Unifying AoI Minimization and Remote Estimation - Optimal Sensor/Controller Coordination with Random Two-way DelayCho-Hsin Tsai, Chih-Chun WangINFOCOM 2020 · 29 citations
- PIPHEN: Physical Interaction Prediction with Hamiltonian Energy NetworksKewei Chen, Yayu Long, Mingsheng ShangAAAI 2026
