Lune

INFOCOM2026Top-tier venue

Global Convergence for Multi-agent Reinforcement Learning in Unreliable Communication Networks

Pengcheng Dai, Lingjie Duan

2026Year

Abstract

Multi-agent reinforcement learning (MARL) has been widely adopted in embodied artificial intelligence (embodied AI) systems, such as cooperative robotics, autonomous swarms, and wireless-enabled intelligent agents, where decision-making relies on local perception, physical interaction, and limited communication. However, most existing MARL algorithms require centralized training with global state-action information, which is often infeasible in real-world embodied settings due to communication unreliability and execution constraints. In this paper, we propose a distributed and communication-efficient MARL algorithm for embodied multi-agent systems. Each agent employs an approximated policy gradient using only its own action and locally available state and reward information. To further reduce communication and computation costs, we adopt a linear function approximation based on κ-hop neighbor states, naturally matching the local perception and interaction range of embodied agents. We prove global convergence with bounded errors induced by unreliable communication. Simulation results show that our method closely approaches centralized performance while achieving up to 99.3% runtime reduction, and outperforms Soft Actor-Critic (SAC) with a 76.0% reduction in computation time. The source code of the proposed method is available at: https://github.com/Pengcheng-Dai/DACA.

Ask about this paper

Ask your agent about it.

Lune has read the top-tier papers around this one, so every answer names the papers it rests on.

Questions to start from

Your agent calls

Lunesearch_papers

Ask in Lune

Free to start. No credit card required.

lune papers get 785edf95-ff4e-47cb-8799-da7c20f25a42

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines