Anytime Detection of Strategic Deviations in Multi-Agent Systems
Etienne Gauthier, Francis Bach, Michael Jordan
Abstract
In many multi-agent systems, agents interact repeatedly and are expected to settle into stable, rational behavior over time. Yet in practice, behavior often drifts, and detecting such deviations in real time remains an open challenge. We introduce a sequential testing framework that monitors whether observed play is consistent with a benchmark of strategic behavior, without assuming a fixed sample size. Our approach builds on the e-value framework for safe anytime-valid inference: by "betting" against the benchmark, we construct a test supermartingale that accumulates evidence whenever observed payoffs systematically violate the expected conditions. For repeated normal-form games, we take equilibrium as the benchmark, yielding a statistically sound, interpretable measure of departure from equilibrium that can be monitored online; our framework unifies the treatment of Nash, correlated, and coarse correlated equilibria, offering finite-time guarantees and a detailed analysis of detection times. We also leverage Benjamini-Hochberg-type procedures to increase detection power in large games while rigorously controlling the false discovery rate. Finally, we extend our method to stochastic games, verifying online whether observed trajectories adhere to a specified target policy, such as a computed equilibrium, broadening the framework's applicability to dynamic, state-dependent settings.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 7b50ff09-ee39-442d-9f24-7b5605669ee9Builds on9
- Algorithmic Collective Action in Machine LearningMoritz Hardt, Eric Mazumdar, Celestine Mendler-Dünner, Tijana ZrnicICML 2023 · 36 citations
- Optimal Best-Arm Identification Methods for Tail-Risk MeasuresShubhada Agrawal, Wouter M. Koolen, Sandeep JunejaNeurIPS 2021 · 34 citations
- Sequential Predictive Two-Sample and Independence TestingAleksandr Podkopaev, Aaditya RamdasNeurIPS 2023 · 29 citations
- Auditing Fairness by BettingBen Chugg, Santiago Cortes-Gomez, Bryan Wilder, Aaditya RamdasNeurIPS 2023 · 29 citations
- Sequential Kernelized Independence TestingAleksandr Podkopaev, Patrick Blöbaum, Shiva Prasad Kasiviswanathan, Aaditya RamdasICML 2023 · 25 citations
Related papers
- Detecting Rewards Deterioration in Episodic Reinforcement LearningIdo Greenberg, Shie MannorICML 2021 · 15 citations
- Learning to Bet for Horizon-Aware Anytime-Valid TestingEge Onur Taga, Samet Oymak, Shubhanshu ShekharICML 2026 · 2 citations
- Sequential Kernel Goodness-of-fit TestingZhengyu Zhou, Weiwei LiuICML 2024 · 1 citation
- A New Framework for Online Testing of Heterogeneous Treatment EffectMiao Yu, Wenbin Lu, Rui SongAAAI 2020 · 10 citations
- Semi-Supervised Hypothesis Testing by Betting on PredictionsYaniv Tenzer, Elad Tolochinksy, Yaniv RomanoICML 2026
