Inverse Game Theory for Stackelberg Games: the Blessing of Bounded Rationality
Jibang Wu, Weiran Shen, Fei Fang, Haifeng Xu
摘要
Optimizing strategic decisions (a.k.a. computing equilibrium) is key to the success of many non-cooperative multi-agent applications. However, in many real-world situations, we may face the exact opposite of this game-theoretic problem -instead of prescribing equilibrium of a given game, we may directly observe the agents' equilibrium behaviors but want to infer the underlying parameters of an unknown game. This research question, also known as inverse game theory, has been studied in multiple recent works in the context of Stackelberg games. Unfortunately, existing works exhibit quite negative results, showing statistical hardness [27, 37] and computational hardness [24, 25, 26] , assuming follower's perfectly rational behaviors. Our work relaxes the perfect rationality agent assumption to the classic quantal response model, a more realistic behavior model of bounded rationality. Interestingly, we show that the smooth property brought by such bounded rationality model actually leads to provably more efficient learning of the follower utility parameters in general Stackelberg games. Systematic empirical experiments on synthesized games confirm our theoretical results and further suggest its robustness beyond the strict quantal response model.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- On Feasible Rewards in Multi-Agent Inverse Reinforcement LearningTill Freihaut, Giorgia RamponiNeurIPS 2025 · 被引用 5 次
- Deviate or Not: Learning Coalition Structures with Multiple-bit Observations in GamesYixuan Even Xu, Zhe Feng, Fei FangAAAI 2025 · 被引用 1 次
- Decoding Rewards in Competitive Games: Inverse Game Theory with Entropy RegularizationJunyi Liao, Zihan Zhu, Ethan X. Fang, Zhuoran Yang 等ICML 2025
- Data-Driven Knowledge-Aware Inference of Private Information in Continuous Double AuctionsLvye Cui, Haoran YuAAAI 2024
- Inferring Heterogeneous Private Valuations from Offline Market Data via Entropic Risk-Sensitive Utility MaximizationXingyu Qian, Haoran YuAAAI 2026
它引用的顶会 Paper2
相关 Paper
- Optimal Rates for Feasible Payoff Set Estimation in GamesAnnalisa Barbara, Riccardo Poiani, Martino Bernasconi, Andrea CelliICML 2026 · 被引用 1 次
- From Behavioral Theories to Econometrics: Inferring Preferences of Human Agents from Data on Repeated InteractionsGali NotiAAAI 2021 · 被引用 7 次
- Stackelberg Learning with Outcome-based PaymentTom Yan, Chicheng ZhangNeurIPS 2025
- Bounded Risk-Sensitive Markov Games: Forward Policy Design and Inverse Reward Learning with Iterative Reasoning and Cumulative Prospect TheoryRan Tian, Liting Sun, Masayoshi TomizukaAAAI 2021 · 被引用 13 次
- Multi-Agent Learning from LearnersMine Melodi Caliskan, Francesco Chini, Setareh MaghsudiICML 2023
