Markov Persuasion Processes: Learning to Persuade From Scratch
Francesco Bacchiocchi, Francesco Emanuele Stradi, Matteo Castiglioni, Alberto Marchesi, Nicola Gatti
摘要
In Bayesian persuasion, an informed sender strategically discloses information to a receiver so as to persuade them to undertake desirable actions. Recently, a growing attention has been devoted to settings in which sender and receivers interact sequentially. Recently, Markov persuasion processes (MPPs) have been introduced to capture sequential scenarios where a sender faces a stream of myopic receivers in a Markovian environment. The MPPs studied so far in the literature suffer from issues that prevent them from being fully operational in practice, e.g., they assume that the sender knows receivers' rewards. We fix such issues by addressing MPPs where the sender has no knowledge about the environment. We design a learning algorithm for the sender, working with partial feedback. We prove that its regret with respect to an optimal information-disclosure policy grows sublinearly in the number of episodes, as it is the case for the loss in persuasiveness cumulated while learning. Moreover, we provide a lower bound for our setting matching the guarantees of our algorithm.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Online Bayesian Persuasion Without a ClueFrancesco Bacchiocchi, Matteo Bollini, Matteo Castiglioni, Alberto Marchesi 等NeurIPS 2024 · 被引用 9 次
- Online Learning in CMDPs: Handling Stochastic and Adversarial ConstraintsFrancesco Emanuele Stradi, Jacopo Germano, Gianmarco Genalti, Matteo Castiglioni 等ICML 2024 · 被引用 7 次
- Taming Adversarial Constraints in CMDPsFrancesco Emanuele Stradi, Anna Lunghi, Matteo Castiglioni, Alberto Marchesi 等NeurIPS 2025 · 被引用 2 次
- Generalized Principal-Agent Problem with a Learning AgentTao Lin, Yiling ChenICLR 2025
- Provably Efficient Algorithm for Best Scoring Rule Identification in Online Principal-Agent Information AcquisitionZichen Wang, Chuanhao Li, Huazheng WangICML 2025
它引用的顶会 Paper14
- Learning Adversarial Markov Decision Processes with Bandit Feedback and Unknown TransitionChi Jin, Tiancheng Jin, Haipeng Luo, Suvrit Sra 等ICML 2020 · 被引用 117 次
- Upper Confidence Primal-Dual Reinforcement Learning for CMDP with Adversarial LossShuang Qiu, Xiaohan Wei, Zhuoran Yang, Jieping Ye 等NeurIPS 2020 · 被引用 65 次
- Signaling in Bayesian Network Congestion Games: the Subtle Power of SymmetryMatteo Castiglioni, Andrea Celli, Alberto Marchesi, Nicola GattiAAAI 2021 · 被引用 44 次
- Multi-Receiver Online Bayesian PersuasionMatteo Castiglioni, Alberto Marchesi, Andrea Celli, Nicola GattiICML 2021 · 被引用 36 次
- Persuading Voters: It's Easy to Whisper, It's Hard to Speak LoudMatteo Castiglioni, Andrea Celli, Nicola GattiAAAI 2020 · 被引用 31 次
相关 Paper
- Optimal Rates and Efficient Algorithms for Online Bayesian PersuasionMartino Bernasconi, Matteo Castiglioni, Andrea Celli, Alberto Marchesi 等ICML 2023 · 被引用 26 次
- Online Bayesian PersuasionMatteo Castiglioni, Andrea Celli, Alberto Marchesi, Nicola GattiNeurIPS 2020 · 被引用 26 次
- Sequential Information Design: Learning to Persuade in the DarkMartino Bernasconi, Matteo Castiglioni, Alberto Marchesi, Nicola Gatti 等NeurIPS 2022 · 被引用 19 次
- Persuading Farsighted Receivers in MDPs: the Power of HonestyMartino Bernasconi, Matteo Castiglioni, Alberto Marchesi, Mirco MuttiNeurIPS 2023 · 被引用 8 次
- Computational Aspects of Bayesian Persuasion under Approximate Best ResponseKunhe Yang, Hanrui ZhangNeurIPS 2024 · 被引用 10 次
