Learning to Play Multi-Follower Bayesian Stackelberg Games
Gerson Personnat, Tao Lin, Safwan Hossain, David C. Parkes
摘要
In a multi-follower Bayesian Stackelberg game, a leader plays a mixed strategy over actions to which followers, each having one of possible private types, best respond. The leader's optimal strategy depends on the distribution of the followers' private types. We study an online learning version of this problem: a leader interacts for rounds with followers with types sampled from an unknown distribution every round. The leader's goal is to minimize regret, defined as the difference between the cumulative utility of the optimal strategy and that of the actually chosen strategies. We design learning algorithms for the leader under different feedback settings. Under type feedback, where the leader observes the followers' types after each round, we design algorithms that achieve regret for independent type distributions and regret for general type distributions. Interestingly, those bounds do not grow with at a polynomial rate. Under action feedback, where the leader only observes the followers' actions, we design algorithms with regret. We also provide a lower bound of , almost matching the type-feedback upper bounds.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Learning in Structured Stackelberg GamesNina Balcan, Kiriaki Fragkia, Keegan HarrisICML 2026 · 被引用 4 次
- Optimally Auditing Adversarial AgentsSanmay Das, Fang-Yi Yu, Yuang ZhangAAAI 2026 · 被引用 1 次
- Learning in Bayesian Stackelberg Games With Unknown Follower's TypesMatteo Bollini, Francesco Bacchiocchi, Samuel Coutts, Matteo Castiglioni 等ICML 2026
它引用的顶会 Paper8
- Optimal Rates and Efficient Algorithms for Online Bayesian PersuasionMartino Bernasconi, Matteo Castiglioni, Andrea Celli, Alberto Marchesi 等ICML 2023 · 被引用 26 次
- Online Bayesian PersuasionMatteo Castiglioni, Andrea Celli, Alberto Marchesi, Nicola GattiNeurIPS 2020 · 被引用 26 次
- Online Learning in Stackelberg Games with an Omniscient FollowerGeng Zhao, Banghua Zhu, Jiantao Jiao, Michael I. JordanICML 2023 · 被引用 23 次
- On the Tractability of Public Persuasion with No ExternalitiesHaifeng XuSODA 2020 · 被引用 22 次
- Computational Aspects of Bayesian Persuasion under Approximate Best ResponseKunhe Yang, Hanrui ZhangNeurIPS 2024 · 被引用 10 次
相关 Paper
- Nearly-Optimal Bandit Learning in Stackelberg Games with Side InformationNina Balcan, Martino Bernasconi, Matteo Castiglioni, Andrea Celli 等ICLR 2026 · 被引用 9 次
- Regret Minimization in Stackelberg Games with Side InformationKeegan Harris, Zhiwei Steven Wu, Maria-Florina BalcanNeurIPS 2024 · 被引用 13 次
- Multi-Receiver Online Bayesian PersuasionMatteo Castiglioni, Alberto Marchesi, Andrea Celli, Nicola GattiICML 2021 · 被引用 36 次
- Learning to Incentivize in Repeated Principal-Agent Problems with Adversarial Agent ArrivalsJunyan Liu, Arnab Maiti, Artin Tajdini, Kevin Jamieson 等ICML 2025
- Learning to Play Sequential Games versus Unknown OpponentsPier Giuseppe Sessa, Ilija Bogunovic, Maryam Kamgarpour, Andreas KrauseNeurIPS 2020 · 被引用 34 次
