Anytime Detection of Strategic Deviations in Multi-Agent Systems
Etienne Gauthier, Francis Bach, Michael Jordan
摘要
In many multi-agent systems, agents interact repeatedly and are expected to settle into stable, rational behavior over time. Yet in practice, behavior often drifts, and detecting such deviations in real time remains an open challenge. We introduce a sequential testing framework that monitors whether observed play is consistent with a benchmark of strategic behavior, without assuming a fixed sample size. Our approach builds on the e-value framework for safe anytime-valid inference: by "betting" against the benchmark, we construct a test supermartingale that accumulates evidence whenever observed payoffs systematically violate the expected conditions. For repeated normal-form games, we take equilibrium as the benchmark, yielding a statistically sound, interpretable measure of departure from equilibrium that can be monitored online; our framework unifies the treatment of Nash, correlated, and coarse correlated equilibria, offering finite-time guarantees and a detailed analysis of detection times. We also leverage Benjamini-Hochberg-type procedures to increase detection power in large games while rigorously controlling the false discovery rate. Finally, we extend our method to stochastic games, verifying online whether observed trajectories adhere to a specified target policy, such as a computed equilibrium, broadening the framework's applicability to dynamic, state-dependent settings.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper9
- Algorithmic Collective Action in Machine LearningMoritz Hardt, Eric Mazumdar, Celestine Mendler-Dünner, Tijana ZrnicICML 2023 · 被引用 36 次
- Optimal Best-Arm Identification Methods for Tail-Risk MeasuresShubhada Agrawal, Wouter M. Koolen, Sandeep JunejaNeurIPS 2021 · 被引用 34 次
- Sequential Predictive Two-Sample and Independence TestingAleksandr Podkopaev, Aaditya RamdasNeurIPS 2023 · 被引用 29 次
- Auditing Fairness by BettingBen Chugg, Santiago Cortes-Gomez, Bryan Wilder, Aaditya RamdasNeurIPS 2023 · 被引用 29 次
- Sequential Kernelized Independence TestingAleksandr Podkopaev, Patrick Blöbaum, Shiva Prasad Kasiviswanathan, Aaditya RamdasICML 2023 · 被引用 25 次
相关 Paper
- Detecting Rewards Deterioration in Episodic Reinforcement LearningIdo Greenberg, Shie MannorICML 2021 · 被引用 15 次
- Learning to Bet for Horizon-Aware Anytime-Valid TestingEge Onur Taga, Samet Oymak, Shubhanshu ShekharICML 2026 · 被引用 2 次
- Sequential Kernel Goodness-of-fit TestingZhengyu Zhou, Weiwei LiuICML 2024 · 被引用 1 次
- A New Framework for Online Testing of Heterogeneous Treatment EffectMiao Yu, Wenbin Lu, Rui SongAAAI 2020 · 被引用 10 次
- Semi-Supervised Hypothesis Testing by Betting on PredictionsYaniv Tenzer, Elad Tolochinksy, Yaniv RomanoICML 2026
