Robust Contextual Combinatorial Multi-Armed Bandits for Unreliable Network Systems
Junkai Wang, Xutong Liu, Jinhang Zuo, Yuedong Xu
摘要
Combinatorial multi-armed bandit (CMAB) is a fundamental framework widely used in networked systems to maximize cumulative rewards under uncertainty. Real-world applications such as federated learning and content delivery network often involve feedback that may be corrupted due to adversarial attacks or network disruptions. In this paper, we study contextual CMAB (MAB) with adversarial corruptions, where feedback for base arms within any selected super arms can be corrupted by an adversary. We focus on-norm smooth reward function and bothand-norm corruption measures, establishing tight regret upper bounds for each scenario. Additionally, we provide the first lower bounds forMAB under corruptions, confirming the optimality of our proposed algorithm. To broaden the applicability, we further extend our algorithm to a more general- MAB setting with probabilistically triggered arms. Empirical validation demonstrates significant improvements across synthetic and real-world datasets, with applications in contextual latency-critic federated learning, user-specific online content delivery and 360° VR video streaming.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- Robust Bandit Learning with Imperfect ContextJianyi Yang, Shaolei RenAAAI 2021 · 被引用 9 次
- Contextual Combinatorial Bandits with Probabilistically Triggered ArmsXutong Liu, Jinhang Zuo, Siwei Wang, John C. S. Lui 等ICML 2023 · 被引用 26 次
- Learning Context-Aware Probabilistic Maximum Coverage Bandits: A Variance-Adaptive ApproachXutong Liu, Jinhang Zuo, Junkai Wang, Zhiyong Wang 等INFOCOM 2024 · 被引用 5 次
- Federated Linear Bandits with Finite Adversarial ActionsLi Fan, Ruida Zhou, Chao Tian, Cong ShenNeurIPS 2023 · 被引用 4 次
- Constraint-Aware Combinatorial Bandits: Theoretical Foundations and Network ApplicationsXiangxiang Dai, Jin Li, Xutong Liu, Anqi Yu 等INFOCOM 2026
