Optimal PAC Multiple Arm Identification with Applications to Crowdsourcing

Yuan Zhou Carnegie Mellon U
Xi Chen Uc Berkeley
Jian Li Tsinghua U

Publication date

September 2015

Abstract

We study the problem of selecting K arms with the highest expected rewards in a stochastic n-armed bandit game. Instead of using existing evaluation metrics (e.g., misidentification probability (Bubeck et al., 2013) or the metric in EXPLORE-K (Kalyanakrishnan & Stone, 2010)), we propose to use the aggregate regret, which is defined as the gap between the average reward of the optimal solution and that of our solution. Besides being a natural metric by itself, we argue that in many applications, such as our motivating example from crowdsourcing, the aggregate regret bound is more suitable. We propose a new PAC algorithm, which, with probability at least 1 − δ, identifies a set of K arms with regret at most . We provide the sample complex...

Extracted data

We use cookies to provide a better user experience.

Data Protection

Optimal PAC Multiple Arm Identification with Applications to Crowdsourcing

Abstract

Extracted data

Optimal PAC Multiple Arm Identification with Applications to Crowdsourcing

Abstract

Extracted data

Related items

Related items