eQMARL: Entangled Quantum Multi-Agent Reinforcement Learning for Distributed Cooperation over Quantum Channels
Alexander C. DeRieux, Walid Saad
摘要
Collaboration is a key challenge in distributed multi-agent reinforcement learning (MARL) environments. Learning frameworks for these decentralized systems must weigh the benefits of explicit player coordination against the communication overhead and computational cost of sharing local observations and environmental data. Quantum computing has sparked a potential synergy between quantum entanglement and cooperation in multi-agent environments, which could enable more efficient distributed collaboration with minimal information sharing. This relationship is largely unexplored, however, as current state-of-the-art quantum MARL (QMARL) implementations rely on classical information sharing rather than entanglement over a quantum channel as a coordination medium. In contrast, in this paper, a novel framework dubbed entangled QMARL (eQMARL) is proposed. The proposed eQMARL is a distributed actor-critic framework that facilitates cooperation over a quantum channel and eliminates local observation sharing via a quantum entangled split critic. Introducing a quantum critic uniquely spread across the agents allows coupling of local observation encoders through entangled input qubits over a quantum channel, which requires no explicit sharing of local observations and reduces classical communication overhead. Further, agent policies are tuned through joint observation-value function estimation via joint quantum measurements, thereby reducing the centralized computational burden. Experimental results show that eQMARL with entanglement converges to a cooperative strategy up to faster and with a higher overall score compared to split classical and fully centralized classical and quantum baselines. The results also show that eQMARL achieves this performance with a constant factor of -times fewer centralized parameters compared to the split classical baseline.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper1
相关 Paper
- Joint Optimization of Circuit Transformation and Qubit Mapping for Distributed Quantum ComputingXiangzhi Zhang, Xu Xu, Yu Liu, Yingling Mao 等INFOCOM 2026 · 被引用 2 次
- Quantum Multi-Agent Meta Reinforcement LearningWon Joon Yun, Jihong Park, Joongheon KimAAAI 2023 · 被引用 50 次
- Solving Continuous Control via Q-learningTim Seyde, Peter Werner, Wilko Schwarting, Igor Gilitschenski 等ICLR 2023 · 被引用 3 次
- PMAC: Personalized Multi-Agent CommunicationXiangrui Meng, Ying TanAAAI 2024 · 被引用 7 次
- Shared Experience Actor-Critic for Multi-Agent Reinforcement LearningFilippos Christianos, Lukas Schäfer, Stefano V. AlbrechtNeurIPS 2020 · 被引用 238 次
