Persuading Farsighted Receivers in MDPs: the Power of Honesty
Martino Bernasconi, Matteo Castiglioni, Alberto Marchesi, Mirco Mutti
摘要
Bayesian persuasion studies the problem faced by an informed sender who strategically discloses information to influence the behavior of an uninformed receiver. Recently, a growing attention has been devoted to settings where the sender and the receiver interact sequentially, in which the receiver's decision-making problem is usually modeled as a Markov decision process (MDP). However, previous works focused on computing optimal information-revelation policies (a.k.a. signaling schemes) under the restrictive assumption that the receiver acts myopically, selecting actions to maximize the one-step utility and disregarding future rewards. This is justified by the fact that, when the receiver is farsighted and thus considers future rewards, finding an optimal Markovian signaling scheme is NP-hard. In this paper, we show that Markovian signaling schemes do not constitute the"right"class of policies. Indeed, differently from most of the MDPs settings, we prove that Markovian signaling schemes are not optimal, and general history-dependent signaling schemes should be considered. Moreover, we also show that history-dependent signaling schemes circumvent the negative complexity results affecting Markovian signaling schemes. Formally, we design an algorithm that computes an optimal and -persuasive history-dependent signaling scheme in time polynomial in 1/ and in the instance size. The crucial challenge is that general history-dependent signaling schemes cannot be represented in polynomial space. Nevertheless, we introduce a convenient subclass of history-dependent signaling schemes, called promise-form, which are as powerful as general history-dependent ones and efficiently representable. Intuitively, promise-form signaling schemes compactly encode histories in the form of honest promises on future receiver's rewards.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Markov Persuasion Processes: Learning to Persuade From ScratchFrancesco Bacchiocchi, Francesco Emanuele Stradi, Matteo Castiglioni, Alberto Marchesi 等NeurIPS 2025 · 被引用 13 次
- Steering the Herd: A Framework for LLM-based Control of Social LearningRaghu Arghal, Kevin He, Shirin Saeedi Bidokhti, Saswati SarkarICLR 2026 · 被引用 1 次
- Stochastic Principal-Agent Problems: Computing and Learning Optimal History-Dependent PoliciesJiarui Gan, Rupak Majumdar, Debmalya Mandal, Goran RadanovicNeurIPS 2025
- Optimal Robust Subsidy Policies for Irrational Agent in Principal-Agent MDPsBowen Hu, Yixin TaoICLR 2026
它引用的顶会 Paper4
- Signaling in Bayesian Network Congestion Games: the Subtle Power of SymmetryMatteo Castiglioni, Andrea Celli, Alberto Marchesi, Nicola GattiAAAI 2021 · 被引用 44 次
- Persuading Voters: It's Easy to Whisper, It's Hard to Speak LoudMatteo Castiglioni, Andrea Celli, Nicola GattiAAAI 2020 · 被引用 31 次
- Bayesian Persuasion in Sequential Decision-MakingJiarui Gan, Rupak Majumdar, Goran Radanovic, Adish SinglaAAAI 2022 · 被引用 30 次
- Sequential Information Design: Learning to Persuade in the DarkMartino Bernasconi, Matteo Castiglioni, Alberto Marchesi, Nicola Gatti 等NeurIPS 2022 · 被引用 19 次
相关 Paper
- Private Bayesian Persuasion with Sequential GamesAndrea Celli, Stefano Coniglio, Nicola GattiAAAI 2020 · 被引用 29 次
- Algorithms for Persuasion with Limited CommunicationRonen Gradwohl, Niklas Hahn, Martin Hoefer, Rann SmorodinskySODA 2021 · 被引用 6 次
- Online Bayesian PersuasionMatteo Castiglioni, Andrea Celli, Alberto Marchesi, Nicola GattiNeurIPS 2020 · 被引用 26 次
- Algorithmic Bayesian Persuasion with Combinatorial ActionsKaito Fujii, Shinsaku SakaueAAAI 2022 · 被引用 3 次
- Computational Aspects of Bayesian Persuasion under Approximate Best ResponseKunhe Yang, Hanrui ZhangNeurIPS 2024 · 被引用 10 次
