Influencing Bandits: Arm Selection for Preference Shaping

Nadkarni, Viraj; Manjunath, D.; Moharir, Sharayu

Computer Science > Machine Learning

arXiv:2403.00036 (cs)

[Submitted on 29 Feb 2024]

Title:Influencing Bandits: Arm Selection for Preference Shaping

Authors:Viraj Nadkarni, D. Manjunath, Sharayu Moharir

View PDF HTML (experimental)

Abstract:We consider a non stationary multi-armed bandit in which the population preferences are positively and negatively reinforced by the observed rewards. The objective of the algorithm is to shape the population preferences to maximize the fraction of the population favouring a predetermined arm. For the case of binary opinions, two types of opinion dynamics are considered -- decreasing elasticity (modeled as a Polya urn with increasing number of balls) and constant elasticity (using the voter model). For the first case, we describe an Explore-then-commit policy and a Thompson sampling policy and analyse the regret for each of these policies. We then show that these algorithms and their analyses carry over to the constant elasticity case. We also describe a Thompson sampling based algorithm for the case when more than two types of opinions are present. Finally, we discuss the case where presence of multiple recommendation systems gives rise to a trade-off between their popularity and opinion shaping objectives.

Comments:	14 pages, 8 figures, 24 references, proofs in appendix
Subjects:	Machine Learning (cs.LG); Artificial Intelligence (cs.AI); Information Retrieval (cs.IR); Systems and Control (eess.SY)
ACM classes:	I.2.6
Cite as:	arXiv:2403.00036 [cs.LG]
	(or arXiv:2403.00036v1 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2403.00036

Submission history

From: D Manjunath [view email]
[v1] Thu, 29 Feb 2024 05:59:27 UTC (1,124 KB)

Computer Science > Machine Learning

Title:Influencing Bandits: Arm Selection for Preference Shaping

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Influencing Bandits: Arm Selection for Preference Shaping

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators