Lune

ICLR2026Top-tier venue

A Near-Optimal Best-of-Both-Worlds Algorithm for Federated Bandits

Zicheng Hu, Zihao Wang, Cheng Chen

2026Year
18Citations
1Top-tier citations

Abstract

This paper studies federated multi-armed bandit (MAB) problems in which multiple agents work together to solve a common MAB problem through a communication network. We focus on the heterogeneous setting in which no single agent can identify the globally best arm using only locally biased observations. In this setting, different agents may select the same arm at the same time step, but receive different rewards. We propose a novel algorithm called FedFTRL for this problem and, to our knowledge, it is the first to achieve near-optimal regret guarantees in both stochastic and adversarial environments. Notably, in the adversarial regime, our algorithm achieves O(T12)O(T^{\frac{1}{2}}) regret, a significant improvement over the state-of-the-art regret of O(T23)O(T^{\frac{2}{3}}) . We also provide empirical evaluations comparing our algorithm with baseline methods, demonstrating the effectiveness of our approach on both synthetic and real-world datasets.

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext 7e4d7dd0-bc23-4668-b03b-ad1bc213cfd1

Cited by top-tier papers1

Ask how each one uses it

Builds on9

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines