Learning Kernelized Contextual Bandits in a Distributed and Asynchronous Environment
Chuanhao Li, Huazheng Wang, Mengdi Wang, Hongning Wang
摘要
Despite the recent advances in communication-efficient distributed bandit learning, most existing solutions are restricted to parametric models, e.g., linear bandits and generalized linear bandits (GLB). In comparison, kernel bandits, which search for non-parametric functions in a reproducing kernel Hilbert space (RKHS), offer higher modeling capacity. But the only existing work in distributed kernel bandits adopts a synchronous communication protocol, which greatly limits its practical use (e.g., every synchronization step requires all clients to participate and wait for data exchange).In this paper, in order to improve the robustness against delays and unavailability of clients that are common in practice, we propose the first asynchronous solution based on approximated kernel regression for distributed kernel bandit learning. A set of effective treatments are developed to ensure approximation quality and communication efficiency. Rigorous theoretical analysis about the regret and communication cost is provided; and extensive empirical evaluations demonstrate the effectiveness of our solution.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper4
- Federated Combinatorial Multi-Agent Multi-Armed BanditsFares Fourati, Mohamed-Slim Alouini, Vaneet AggarwalICML 2024 · 被引用 10 次
- Incentivized Communication for Federated BanditsZhepei Wei, Chuanhao Li, Haifeng Xu, Hongning WangNeurIPS 2023 · 被引用 4 次
- Communication-Efficient Federated Non-Linear Bandit OptimizationChuanhao Li, Chong Liu, Yu-Xiang WangICLR 2024 · 被引用 2 次
- Incentivized Truthful Communication for Federated BanditsZhepei Wei, Chuanhao Li, Tianze Ren, Haifeng Xu 等ICLR 2024 · 被引用 2 次
相关 Paper
- Communication Efficient Distributed Learning for Kernelized Contextual BanditsChuanhao Li, Huazheng Wang, Mengdi Wang, Hongning WangNeurIPS 2022 · 被引用 19 次
- Kernel Methods for Cooperative Multi-Agent Contextual BanditsAbhimanyu Dubey, Alex 'Sandy' PentlandICML 2020 · 被引用 32 次
- Approximation Theory Based Methods for RKHS BanditsSho Takemori, Masahiro SatoICML 2021 · 被引用 3 次
- Communication Efficient Federated Learning for Generalized Linear BanditsChuanhao Li, Hongning WangNeurIPS 2022 · 被引用 19 次
- Personalized Online Federated Learning with Multiple KernelsPouya M. Ghari, Yanning ShenNeurIPS 2022 · 被引用 20 次
