Bayesian Persuasion in Sequential Decision-Making
Jiarui Gan, Rupak Majumdar, Goran Radanovic, Adish Singla
摘要
We study a dynamic model of Bayesian persuasion in sequential decision-making settings. An informed principal observes an external parameter of the world and advises an uninformed agent about actions to take over time. The agent takes actions in each time step based on the current state, the principal's advice/signal, and beliefs about the external parameter. The action of the agent updates the state according to a stochastic process. The model arises naturally in many applications, e.g., an app (the principal) can advice the user (the agent) on possible choices between actions based on additional real-time information the app has. We study the problem of designing a signaling strategy from the principal's point of view. We show that the principal has an optimal strategy against a myopic agent, who only optimizes their rewards locally, and the optimal strategy can be computed in polynomial time. In contrast, it is NP-hard to approximate an optimal policy against a far-sighted agent. Further, we show that if the principal has the power to threaten the agent by not providing future signals, then we can efficiently design a threat-based strategy. This strategy guarantees the principal's payoff as if playing against an agent who is far-sighted but myopic to future signals.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- Information Design in Multi-Agent Reinforcement LearningYue Lin, Wenhao Li, Hongyuan Zha, Baoxiang WangNeurIPS 2023 · 被引用 25 次
- Sequential Information Design: Learning to Persuade in the DarkMartino Bernasconi, Matteo Castiglioni, Alberto Marchesi, Nicola Gatti 等NeurIPS 2022 · 被引用 19 次
- Polynomial-Time Optimal Equilibria with a Mediator in Extensive-Form GamesBrian Hu Zhang, Tuomas SandholmNeurIPS 2022 · 被引用 15 次
- Markov Persuasion Processes: Learning to Persuade From ScratchFrancesco Bacchiocchi, Francesco Emanuele Stradi, Matteo Castiglioni, Alberto Marchesi 等NeurIPS 2025 · 被引用 13 次
- Persuading Farsighted Receivers in MDPs: the Power of HonestyMartino Bernasconi, Matteo Castiglioni, Alberto Marchesi, Mirco MuttiNeurIPS 2023 · 被引用 8 次
它引用的顶会 Paper6
- Policy Teaching via Environment Poisoning: Training-time Adversarial Attacks against Reinforcement LearningAmin Rakhsha, Goran Radanovic, Rati Devidze, Xiaojin Zhu 等ICML 2020 · 被引用 145 次
- Multi-Receiver Online Bayesian PersuasionMatteo Castiglioni, Alberto Marchesi, Andrea Celli, Nicola GattiICML 2021 · 被引用 36 次
- Persuading Voters: It's Easy to Whisper, It's Hard to Speak LoudMatteo Castiglioni, Andrea Celli, Nicola GattiAAAI 2020 · 被引用 31 次
- Private Bayesian Persuasion with Sequential GamesAndrea Celli, Stefano Coniglio, Nicola GattiAAAI 2020 · 被引用 29 次
- Online Bayesian PersuasionMatteo Castiglioni, Andrea Celli, Alberto Marchesi, Nicola GattiNeurIPS 2020 · 被引用 26 次
相关 Paper
- Bayesian Persuasion with Externalities: Exploiting Agent TypesJonathan Shaki, Jiarui Gan, Sarit KrausAAAI 2025
- Computational Aspects of Bayesian Persuasion under Approximate Best ResponseKunhe Yang, Hanrui ZhangNeurIPS 2024 · 被引用 10 次
- Signaling in Bayesian Network Congestion Games: the Subtle Power of SymmetryMatteo Castiglioni, Andrea Celli, Alberto Marchesi, Nicola GattiAAAI 2021 · 被引用 44 次
- Automated Dynamic Mechanism DesignHanrui Zhang, Vincent ConitzerNeurIPS 2021 · 被引用 18 次
- Algorithms for Persuasion with Limited CommunicationRonen Gradwohl, Niklas Hahn, Martin Hoefer, Rann SmorodinskySODA 2021 · 被引用 6 次
