R2-B2: Recursive Reasoning-Based Bayesian Optimization for No-Regret Learning in Games
Zhongxiang Dai, Yizhou Chen, Bryan Kian Hsiang Low, Patrick Jaillet, Teck-Hua Ho
摘要
This paper presents a recursive reasoning formalism of Bayesian optimization (BO) to model the reasoning process in the interactions between boundedly rational, self-interested agents with unknown, complex, and costly-to-evaluate payoff functions in repeated games, which we call Recursive Reasoning-Based BO (R2-B2). Our R2-B2 algorithm is general in that it does not constrain the relationship among the payoff functions of different agents and can thus be applied to various types of games such as constant-sum, general-sum, and common-payoff games. We prove that by reasoning at level 2 or more and at one level higher than the other agents, our R2-B2 agent can achieve faster asymptotic convergence to no regret than that without utilizing recursive reasoning. We also propose a computationally cheaper variant of R2-B2 called R2-B2-Lite at the expense of a weaker convergence guarantee. The performance and generality of our R2-B2 algorithm are empirically demonstrated using synthetic games, adversarial machine learning, and multi-agent reinforcement learning.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Federated Bayesian Optimization via Thompson SamplingZhongxiang Dai, Bryan Kian Hsiang Low, Patrick JailletNeurIPS 2020 · 被引用 144 次
- Differentially Private Federated Bayesian Optimization with Distributed ExplorationZhongxiang Dai, Bryan Kian Hsiang Low, Patrick JailletNeurIPS 2021 · 被引用 64 次
- Sample-Then-Optimize Batch Neural Thompson SamplingZhongxiang Dai, Yao Shu, Bryan Kian Hsiang Low, Patrick JailletNeurIPS 2022 · 被引用 33 次
- Efficient Distributionally Robust Bayesian Optimization with Worst-case SensitivitySebastian Shenghong Tay, Chuan Sheng Foo, Daisuke Urano, Richalynn Leong 等ICML 2022 · 被引用 20 次
- Bayesian Optimization under Stochastic Delayed FeedbackArun Verma, Zhongxiang Dai, Bryan Kian Hsiang LowICML 2022 · 被引用 15 次
它引用的顶会 Paper4
- MagNet: A Two-Pronged Defense against Adversarial ExamplesDongyu Meng, Hao ChenCCS 2017 · 被引用 1,295 次
- BayesOpt Adversarial AttackBinxin Ru, Adam D. Cobb, Arno Blaas, Yarin GalICLR 2020 · 被引用 85 次
- Private Outsourced Bayesian OptimizationDmitrii Kharkovskii, Zhongxiang Dai, Bryan Kian Hsiang LowICML 2020 · 被引用 25 次
- Scalable Variational Bayesian Kernel Selection for Sparse Gaussian Process RegressionTong Teng, Jie Chen, Yehong Zhang, Bryan Kian Hsiang LowAAAI 2020 · 被引用 24 次
相关 Paper
- Recursive Reasoning Graph for Multi-Agent Reinforcement LearningXiaobai Ma, David Isele, Jayesh K. Gupta, Kikuo Fujimura 等AAAI 2022 · 被引用 8 次
- Generalized Principal-Agent Problem with a Learning AgentTao Lin, Yiling ChenICLR 2025
- Adversarial Causal Bayesian OptimizationScott Sussex, Pier Giuseppe Sessa, Anastasia Makarova, Andreas KrauseICLR 2024 · 被引用 5 次
- Collaborative Bayesian Optimization with Fair RegretRachael Hwee Ling Sim, Yehong Zhang, Bryan Kian Hsiang Low, Patrick JailletICML 2021 · 被引用 26 次
- Policy Optimization for Markov Games: Unified Framework and Faster ConvergenceRunyu Zhang, Qinghua Liu, Huan Wang, Caiming Xiong 等NeurIPS 2022 · 被引用 32 次
