Lune

ICML2022Top-tier venue

Feature and Parameter Selection in Stochastic Linear Bandits

Ahmadreza Moradipari, Berkay Turan, Yasin Abbasi-Yadkori, Mahnoosh Alizadeh, Mohammad Ghavamzadeh

2022Year
6Citations
2Top-tier citations

Abstract

We study two model selection settings in stochastic linear bandits (LB). In the first setting, which we refer to as feature selection, the expected reward of the LB problem is in the linear span of at least one of MM feature maps (models). In the second setting, the reward parameter of the LB problem is arbitrarily selected from MM models represented as (possibly) overlapping balls in Rd\mathbb R^d. However, the agent only has access to misspecified models, i.e., estimates of the centers and radii of the balls. We refer to this setting as parameter selection. For each setting, we develop and analyze a computationally efficient algorithm that is based on a reduction from bandits to full-information problems. This allows us to obtain regret bounds that are not worse (up to a log⁡M\sqrt{\log M} factor) than the case where the true model is known. This is the best-reported dependence on the number of models MM in these settings. Finally, we empirically show the effectiveness of our algorithms using synthetic and real-world experiments.

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext b098f405-0ae4-4f11-a9cd-e7afe7399551

Cited by top-tier papers2

Ask how each one uses it

Builds on8

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines