Cascading Hybrid Bandits: Online Learning to Rank for Relevance and Diversity

Li, Chang; Feng, Haoyun; de Rijke, Maarten

doi:10.1145/3383313.3412245

Computer Science > Machine Learning

arXiv:1912.00508 (cs)

[Submitted on 1 Dec 2019 (v1), last revised 12 Aug 2020 (this version, v3)]

Title:Cascading Hybrid Bandits: Online Learning to Rank for Relevance and Diversity

Authors:Chang Li, Haoyun Feng, Maarten de Rijke

View PDF

Abstract:Relevance ranking and result diversification are two core areas in modern recommender systems. Relevance ranking aims at building a ranked list sorted in decreasing order of item relevance, while result diversification focuses on generating a ranked list of items that covers a broad range of topics. In this paper, we study an online learning setting that aims to recommend a ranked list with $K$ items that maximizes the ranking utility, i.e., a list whose items are relevant and whose topics are diverse. We formulate it as the cascade hybrid bandits (CHB) problem. CHB assumes the cascading user behavior, where a user browses the displayed list from top to bottom, clicks the first attractive item, and stops browsing the rest. We propose a hybrid contextual bandit approach, called CascadeHybrid, for solving this problem. CascadeHybrid models item relevance and topical diversity using two independent functions and simultaneously learns those functions from user click feedback. We conduct experiments to evaluate CascadeHybrid on two real-world recommendation datasets: MovieLens and Yahoo music datasets. Our experimental results show that CascadeHybrid outperforms the baselines. In addition, we prove theoretical guarantees on the $n$-step performance demonstrating the soundness of CascadeHybrid.

Subjects:	Machine Learning (cs.LG); Information Retrieval (cs.IR); Machine Learning (stat.ML)
Cite as:	arXiv:1912.00508 [cs.LG]
	(or arXiv:1912.00508v3 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.1912.00508
Related DOI:	https://doi.org/10.1145/3383313.3412245

Submission history

From: Chang Li [view email]
[v1] Sun, 1 Dec 2019 22:03:18 UTC (4,905 KB)
[v2] Tue, 3 Dec 2019 16:57:55 UTC (4,905 KB)
[v3] Wed, 12 Aug 2020 06:46:20 UTC (18,902 KB)

Computer Science > Machine Learning

Title:Cascading Hybrid Bandits: Online Learning to Rank for Relevance and Diversity

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Cascading Hybrid Bandits: Online Learning to Rank for Relevance and Diversity

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators