Lune

NeurIPS2025Top-tier venue

Multi-agent KTO: Enhancing Strategic Interactions of Large Language Model in Language Game

Rong Ye, Yongxin Zhang, Yikai Zhang, Haoyu Kuang, Peng Sun, Zhongyu Wei

2025Year

Abstract

Achieving Artificial General Intelligence (AGI) requires AI agents that can not only make strategic decisions but also engage in flexible and meaningful communication. Inspired by Wittgenstein’s language game theory, we propose that language agents can learn through in-context interaction rather than traditional multi-stage frameworks that separate decision-making from language expression. Using Werewolf , a social deduction game that tests language understanding, strategic interaction, and adaptability, as a test bed, we develop the Multi-agent Kahneman-Tversky’s Optimization (MaKTO). MaKTO engages diverse models in extensive gameplay to generate unpaired desirable and unacceptable responses, then employs KTO to refine the model’s decision-making process. In 9-player Werewolf games, MaKTO achieves a 61% average win rate across various models, outperforming GPT-4o and two-stage RL agents by relative improvements of 23.0% and 10.9%, respectively. Notably, MaKTO also demonstrates human-like performance, winning 60% against expert players and showing only 48.9% de-tectability in Turing-style blind tests. Code and data are available at project page https://reneeye.github.io/MaKTO.html .

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext 6f67c8b7-e76f-4dea-9901-76402c47209e

Builds on18

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines