Almost Optimal Algorithms for Linear Stochastic Bandits with Heavy-Tailed Payoffs

Shao, Han; Yu, Xiaotian; King, Irwin; Lyu, Michael R.

Computer Science > Machine Learning

arXiv:1810.10895 (cs)

[Submitted on 25 Oct 2018 (v1), last revised 11 Nov 2018 (this version, v2)]

Title:Almost Optimal Algorithms for Linear Stochastic Bandits with Heavy-Tailed Payoffs

Authors:Han Shao, Xiaotian Yu, Irwin King, Michael R. Lyu

View PDF

Abstract:In linear stochastic bandits, it is commonly assumed that payoffs are with sub-Gaussian noises. In this paper, under a weaker assumption on noises, we study the problem of \underline{lin}ear stochastic {\underline b}andits with h{\underline e}avy-{\underline t}ailed payoffs (LinBET), where the distributions have finite moments of order $1+\epsilon$, for some $\epsilon\in (0,1]$. We rigorously analyze the regret lower bound of LinBET as $\Omega(T^{\frac{1}{1+\epsilon}})$, implying that finite moments of order 2 (i.e., finite variances) yield the bound of $\Omega(\sqrt{T})$, with $T$ being the total number of rounds to play bandits. The provided lower bound also indicates that the state-of-the-art algorithms for LinBET are far from optimal. By adopting median of means with a well-designed allocation of decisions and truncation based on historical information, we develop two novel bandit algorithms, where the regret upper bounds match the lower bound up to polylogarithmic factors. To the best of our knowledge, we are the first to solve LinBET optimally in the sense of the polynomial order on $T$. Our proposed algorithms are evaluated based on synthetic datasets, and outperform the state-of-the-art results.

Subjects:	Machine Learning (cs.LG); Machine Learning (stat.ML)
Cite as:	arXiv:1810.10895 [cs.LG]
	(or arXiv:1810.10895v2 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.1810.10895

Submission history

From: Xiaotian Yu [view email]
[v1] Thu, 25 Oct 2018 14:29:02 UTC (284 KB)
[v2] Sun, 11 Nov 2018 13:47:50 UTC (294 KB)

Full-text links:

Access Paper:

view license

Current browse context:

cs.LG

< prev | next >

new | recent | 2018-10

Change to browse by:

cs
stat
stat.ML

References & Citations

DBLP - CS Bibliography

listing | bibtex

Han Shao
Xiaotian Yu
Irwin King
Michael R. Lyu

export BibTeX citation

Computer Science > Machine Learning

Title:Almost Optimal Algorithms for Linear Stochastic Bandits with Heavy-Tailed Payoffs

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Almost Optimal Algorithms for Linear Stochastic Bandits with Heavy-Tailed Payoffs

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators