Near-Optimal Online Learning with Non-Stochastic and Unbounded Erroneous Feedback
Dacheng Wen, Yupeng Li, Francis C. M. Lau, Tian Wang, Yang Chen
摘要
Online learning is a foundational machine learning paradigm in both academia and industry. Most existing online learning techniques are designed for idealized scenarios where the feedback on the decision costs observed by the learner is assumed reliable, i.e., the same as the ground truth. Many recent efforts that attempted to investigate erroneous feedbacks considered errors with unrealistic settings, such as restrictive stochasticity and/or error bounds (e.g., bounded corruption magnitudes or budget of deviations from the ground truths). In this work, we consider a novel and challenging problem of full-information online learning in the presence of feedback with non-stochastic and unbounded errors. According to our analysis, existing representative techniques suffer unbounded regret when applied to our problem. To tackle such erroneous feedback, we propose a robust online learning strategy with a tailored FTRL-like decision-making approach based on a coordinate-wise trimmed sum mechanism, which we prove can achieve a near-optimal, sublinear regret bound of under certain justified conditions. We compare our solution against four representative approaches by evaluating them in three exemplary networking applications. The results not only corroborate our theoretical analysis but also clearly demonstrate the robustness of our algorithm in comparison to the baselines.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- Unconstrained Robust Online Convex OptimizationJiujia Zhang, Ashok CutkoskyICML 2025
- Augment Online Linear Optimization with Arbitrarily Bad Machine-Learned PredictionsDacheng Wen, Yupeng Li, Francis C. M. LauINFOCOM 2024 · 被引用 5 次
- No-Regret Learning Under Adversarial Resource Constraints: A Spending Plan Is All You Need!Francesco Emanuele Stradi, Matteo Castiglioni, Alberto Marchesi, Nicola Gatti 等NeurIPS 2025 · 被引用 7 次
- Understanding the Role of Feedback in Online Learning with Switching CostsDuo Cheng, Xingyu Zhou, Bo JiICML 2023 · 被引用 6 次
- Robust Online Learning against Malicious Manipulation with Application to Network Flow ClassificationYupeng Li, Ben Liang, Ali TizghadamINFOCOM 2021 · 被引用 13 次
