The Influence of Shape Constraints on the Thresholding Bandit Problem

Cheshire, James; Menard, Pierre; Carpentier, Alexandra

Full-text links:

Download:

Current browse context:

cs.LG

< prev | next >

new | recent | 2006

Computer Science > Machine Learning

Title: The Influence of Shape Constraints on the Thresholding Bandit Problem

Authors: James Cheshire, Pierre Menard, Alexandra Carpentier

(Submitted on 17 Jun 2020 (v1), last revised 23 Feb 2021 (this version, v3))

Abstract: We investigate the stochastic Thresholding Bandit problem (TBP) under several shape constraints. On top of (i) the vanilla, unstructured TBP, we consider the case where (ii) the sequence of arm's means $(\mu_k)_k$ is monotonically increasing MTBP, (iii) the case where $(\mu_k)_k$ is unimodal UTBP and (iv) the case where $(\mu_k)_k$ is concave CTBP. In the TBP problem the aim is to output, at the end of the sequential game, the set of arms whose means are above a given threshold. The regret is the highest gap between a misclassified arm and the threshold. In the fixed budget setting, we provide problem independent minimax rates for the expected regret in all settings, as well as associated algorithms. We prove that the minimax rates for the regret are (i) $\sqrt{\log(K)K/T}$ for TBP, (ii) $\sqrt{\log(K)/T}$ for MTBP, (iii) $\sqrt{K/T}$ for UTBP and (iv) $\sqrt{\log\log K/T}$ for CTBP, where $K$ is the number of arms and $T$ is the budget. These rates demonstrate that the dependence on $K$ of the minimax regret varies significantly depending on the shape constraint. This highlights the fact that the shape constraints modify fundamentally the nature of the TBP.

Subjects:	Machine Learning (cs.LG); Machine Learning (stat.ML)
Cite as:	arXiv:2006.10006 [cs.LG]
	(or arXiv:2006.10006v3 [cs.LG] for this version)

Submission history

From: James Cheshire [view email]
[v1] Wed, 17 Jun 2020 17:12:32 GMT (61kb)
[v2] Tue, 12 Jan 2021 15:46:56 GMT (61kb)
[v3] Tue, 23 Feb 2021 09:27:50 GMT (61kb)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> cs > arXiv:2006.10006

Download:

Current browse context:

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

Computer Science > Machine Learning

Title: The Influence of Shape Constraints on the Thresholding Bandit Problem

Submission history