Lune

NeurIPS2025顶会

Non-Uniform Multiclass Learning with Bandit Feedback

Steve Hanneke, Amirreza Shaeiri, Hongao Wang

2025年份
1被引次数

摘要

We study the problem of multiclass learning with bandit feedback in both the i.i.d. batch and adversarial online models. In the uniform learning framework, it is well known that no hypothesis class H is learnable in either model when the effective number of labels is unbounded. In contrast, within the universal learning framework, recent works by Hanneke et al. [2025b] and Hanneke et al. [2025a] have established surprising exact equivalences between learnability under bandit feedback and full supervision in both the i.i.d. batch and adversarial online models, respectively. This raises a natural question: What happens in the nonuniform learning framework, which lies between the uniform and universal learning frameworks? Our contributions are twofold: (1) We provide a combinatorial characterization of learnable hypothesis classes in both models, in the realizable and agnostic settings, within the non-uniform learning framework. Notably, this includes elementary and natural hypothesis classes, such as a countably infinite collection of constant functions over some domain that is learnable in both models. (2) We construct a hypothesis class that is non-uniformly learnable under full supervision in the adversarial online model (and thus also in the i.i.d. batch model), but not non-uniformly learnable under bandit feedback in the i.i.d. batch model (and thus also not in the adversarial online model). This serves as our main novel technical contribution that reveals a fundamental distinction between the non-uniform and universal learning frameworks.

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

它引用的顶会 Paper6

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖