End-to-End Game-Focused Learning of Adversary Behavior in Security Games
Andrew Perrault, Bryan Wilder, Eric Ewing, Aditya Mate, Bistra Dilkina, Milind Tambe
摘要
Stackelberg security games are a critical tool for maximizing the utility of limited defense resources to protect important targets from an intelligent adversary. Motivated by green security, where the defender may only observe an adversary's response to defense on a limited set of targets, we study the problem of learning a defense that generalizes well to a new set of targets with novel feature values and combinations. Traditionally, this problem has been addressed via a two-stage approach where an adversary model is trained to maximize predictive accuracy without considering the defender's optimization problem. We develop an end-to-end game-focused approach, where the adversary model is trained to maximize a surrogate for the defender's expected utility. We show both in theory and experimental results that our game-focused approach achieves higher defender expected utility than the two-stage alternative when there is limited data.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Learning MDPs from Features: Predict-Then-Optimize for Sequential Decision Making by Reinforcement LearningKai Wang, Sanket Shah, Haipeng Chen, Andrew Perrault 等NeurIPS 2021 · 被引用 44 次
- Automatically Learning Compact Quality-aware Surrogates for Optimization ProblemsKai Wang, Bryan Wilder, Andrew Perrault, Milind TambeNeurIPS 2020 · 被引用 37 次
- End-to-end Stochastic Optimization with Energy-based ModelLingkai Kong, Jiaming Cui, Yuchen Zhuang, Rui Feng 等NeurIPS 2022 · 被引用 33 次
- Scalable Decision-Focused Learning in Restless Multi-Armed Bandits with Application to Maternal and Child HealthKai Wang, Shresth Verma, Aditya Mate, Sanket Shah 等AAAI 2023 · 被引用 19 次
- Choices Are Not Independent: Stackelberg Security Games with Nested Quantal Response ModelsTien Mai, Arunesh SinhaAAAI 2022 · 被引用 4 次
相关 Paper
- Fast Algorithms for Stackelberg Prediction Game with Least Squares LossJiali Wang, He Chen, Rujun Jiang, Xudong Li 等ICML 2021 · 被引用 23 次
- Learning to Play Sequential Games versus Unknown OpponentsPier Giuseppe Sessa, Ilija Bogunovic, Maryam Kamgarpour, Andreas KrauseNeurIPS 2020 · 被引用 34 次
- Learning in Structured Stackelberg GamesNina Balcan, Kiriaki Fragkia, Keegan HarrisICML 2026 · 被引用 4 次
- When Can the Defender Effectively Deceive Attackers in Security Games?Thanh Nguyen, Haifeng XuAAAI 2022 · 被引用 4 次
- Security Games with Layered Defenses: Adaptive Adversaries and Gittins IndicesChun Kai Ling, Jakub Cerný, Chin Hui Han, Garud Iyengar 等AAAI 2026
