Safe and Robust Subgame Exploitation in Imperfect Information Games
Zhenxing Ge, Zheng Xu, Tianyu Ding, Linjian Meng, Bo An, Wenbin Li, Yang Gao
Abstract
Opponent exploitation is an important task for players to exploit the weaknesses of others in games. Existing approaches mainly focus on balancing between exploitation and exploitability but are often vulnerable to modeling errors and deceptive adversaries. To address this problem, our paper offers a novel perspective on the safety of opponent exploitation, named Adaptation Safety. This concept leverages the insight that strategies, even those not explicitly aimed at opponent exploitation, may inherently be exploitable due to computational complexities, rendering traditional safety overly rigorous. In contrast, adaptation safety requires that the strategy should not be more exploitable than it would be in scenarios where opponent exploitation is not considered. Building on such adaptation safety, we further propose an Opponent eXploitation Search (OX-Search) framework by incorporating real-time search techniques for efficient online opponent exploitation. Moreover, we provide theoretical analyses to show the adaptation safety and robust exploitation of OX-Search, even with inaccurate opponent models. Empirical evaluations in popular poker games demonstrate OX-Search's superiority in both exploitability and exploitation compared to previous methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d527e6d4-f188-4e0e-bf14-29ed65f002f0Cited by top-tier papers2
- Best of Both Worlds: Regret Minimization versus Minimax PlayAdrian Müller, Jon Schneider, Stratis Skoulakis, Luca Viano et al.ICML 2025
- Safety Game: Inference-Time Alignment of Black-Box LLMs via Constrained OptimizationTuan Nguyen, Long Tran-ThanhICML 2026
Builds on7
- Improving Policies via Search in Cooperative Partially Observable GamesAdam Lerer, Hengyuan Hu, Jakob N. Foerster, Noam BrownAAAI 2020 · 87 citations
- Model-Based Opponent ModelingXiaopeng Yu, Jiechuan Jiang, Wanpeng Zhang, Haobin Jiang et al.NeurIPS 2022 · 56 citations
- Subgame solving without common knowledgeBrian Hu Zhang, Tuomas SandholmNeurIPS 2021 · 21 citations
- Greedy when Sure and Conservative when Uncertain about the OpponentsHaobo Fu, Ye Tian, Hongxiang Yu, Weiming Liu et al.ICML 2022 · 12 citations
- Safe Opponent-Exploitation Subgame RefinementMingyang Liu, Chengjie Wu, Qihan Liu, Yansen Jing et al.NeurIPS 2022 · 9 citations
Related papers
- Opponent-Limited Online Search for Imperfect Information GamesWeiming Liu, Haobo Fu, Qiang Fu, Wei YangICML 2023 · 7 citations
- Safe Learning in Tree-Form Sequential Decision Making: Handling Hard and Soft ConstraintsMartino Bernasconi, Federico Cacciamani, Matteo Castiglioni, Alberto Marchesi et al.ICML 2022 · 10 citations
- Small Nash Equilibrium Certificates in Very Large GamesBrian Hu Zhang, Tuomas SandholmNeurIPS 2020 · 6 citations
- Exploiting Opponents Under Utility Constraints in Sequential GamesMartino Bernasconi de Luca, Federico Cacciamani, Simone Fioravanti, Nicola Gatti et al.NeurIPS 2021 · 10 citations
- MaMa: A Game-Theoretic Approach for Designing Safe Agentic SystemsJonathan Nöther, Adish Singla, Goran RadanovicICML 2026
