Sparsity, variance and curvature in multi-armed bandits

Bubeck, Sébastien; Cohen, Michael B.; Li, Yuanzhi

Computer Science > Machine Learning

arXiv:1711.01037 (cs)

[Submitted on 3 Nov 2017]

Title:Sparsity, variance and curvature in multi-armed bandits

Authors:Sébastien Bubeck, Michael B. Cohen, Yuanzhi Li

View PDF

Abstract:In (online) learning theory the concepts of sparsity, variance and curvature are well-understood and are routinely used to obtain refined regret and generalization bounds. In this paper we further our understanding of these concepts in the more challenging limited feedback scenario. We consider the adversarial multi-armed bandit and linear bandit settings and solve several open problems pertaining to the existence of algorithms with favorable regret bounds under the following assumptions: (i) sparsity of the individual losses, (ii) small variation of the loss sequence, and (iii) curvature of the action set. Specifically we show that (i) for $s$-sparse losses one can obtain $\tilde{O}(\sqrt{s T})$-regret (solving an open problem by Kwon and Perchet), (ii) for loss sequences with variation bounded by $Q$ one can obtain $\tilde{O}(\sqrt{Q})$-regret (solving an open problem by Kale and Hazan), and (iii) for linear bandit on an $\ell_p^n$ ball one can obtain $\tilde{O}(\sqrt{n T})$-regret for $p \in [1,2]$ and one has $\tilde{\Omega}(n \sqrt{T})$-regret for $p>2$ (solving an open problem by Bubeck, Cesa-Bianchi and Kakade). A key new insight to obtain these results is to use regularizers satisfying more refined conditions than general self-concordance

Comments:	18 pages
Subjects:	Machine Learning (cs.LG)
Cite as:	arXiv:1711.01037 [cs.LG]
	(or arXiv:1711.01037v1 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.1711.01037

Submission history

From: Sebastien Bubeck [view email]
[v1] Fri, 3 Nov 2017 06:46:45 UTC (21 KB)

Full-text links:

Access Paper:

view license

Current browse context:

cs.LG

< prev | next >

new | recent | 2017-11

Change to browse by:

References & Citations

DBLP - CS Bibliography

listing | bibtex

Sébastien Bubeck
Michael B. Cohen
Yuanzhi Li

export BibTeX citation

Computer Science > Machine Learning

Title:Sparsity, variance and curvature in multi-armed bandits

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Sparsity, variance and curvature in multi-armed bandits

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators