Partner Selection for the Emergence of Cooperation in Multi-Agent Systems Using Reinforcement Learning
Nicolas Anastassacos, Stephen Hailes, Mirco Musolesi
摘要
Social dilemmas have been widely studied to explain how humans are able to cooperate in society. Considerable effort has been invested in designing artificial agents for social dilemmas that incorporate explicit agent motivations that are chosen to favor coordinated or cooperative responses. The prevalence of this general approach points towards the importance of achieving an understanding of both an agent's internal design and external environment dynamics that facilitate cooperative behavior. In this paper, we investigate how partner selection can promote cooperative behavior between agents who are trained to maximize a purely selfish objective function. Our experiments reveal that agents trained with this dynamic learn a strategy that retaliates against defectors while promoting cooperation with other agents resulting in a prosocial society.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Beyond Mandatory Federations: Balancing Egoism, Utilitarianism and Egalitarianism in Mixed-Motive GamesShaokang Dong, Chao Li, Shangdong Yang, Hongye Cao 等AAAI 2025 · 被引用 1 次
- Moral Alignment for LLM AgentsElizaveta Tennant, Stephen Hailes, Mirco MusolesiICLR 2025
- Learning to Cooperate with Minimal ObservabilityChin-wing Leung, Paolo Turrini, Fernando P. Santos, Mirco MusolesiAAAI 2026
相关 Paper
- Reciprocal Reward Influence Encourages Cooperation From Self-Interested AgentsJohn L. Zhou, Weizhe Hong, Jonathan C. KaoNeurIPS 2024 · 被引用 5 次
- Self-Play Q-Learners Can Provably Collude in the Iterated Prisoner's DilemmaQuentin Bertrand, Juan Agustin Duque, Emilio Calvano, Gauthier GidelICML 2025
- Learning to Incentivize Other Learning AgentsJiachen Yang, Ang Li, Mehrdad Farajtabar, Peter Sunehag 等NeurIPS 2020 · 被引用 105 次
- Emergence of Punishment in Social Dilemma with Environmental FeedbackZhen Wang, Zhao Song, Chen Shen, Shuyue HuAAAI 2023 · 被引用 41 次
- Unsupervised Partner Design Enables Robust Ad-hoc TeamworkConstantin Ruhdorfer, Matteo Bortoletto, Victor Oei, Anna Penzkofer 等ICML 2026 · 被引用 3 次
