Adaptive Decision Making via Entropy Minimization

Allahverdyan, Armen E.; Galstyan, Aram; Abbas, Ali E.; Struzik, Zbigniew R.

Physics > Data Analysis, Statistics and Probability

arXiv:1803.06638 (physics)

[Submitted on 18 Mar 2018 (v1), last revised 1 Dec 2018 (this version, v2)]

Title:Adaptive Decision Making via Entropy Minimization

Authors:Armen E. Allahverdyan, Aram Galstyan, Ali E. Abbas, Zbigniew R. Struzik

View PDF

Abstract:An agent choosing between various actions tends to take the one with the lowest cost. But this choice is arguably too rigid (not adaptive) to be useful in complex situations, e.g., where exploration-exploitation trade-off is relevant in creative task solving or when stated preferences differ from revealed ones. Here we study an agent who is willing to sacrifice a fixed amount of expected utility for adaptation. How can/ought our agent choose an optimal (in a technical sense) mixed action? We explore consequences of making this choice via entropy minimization, which is argued to be a specific example of risk-aversion. This recovers the $\epsilon$-greedy probabilities known in reinforcement learning. We show that the entropy minimization leads to rudimentary forms of intelligent behavior: (i) the agent assigns a non-negligible probability to costly events; but (ii) chooses with a sizable probability the action related to less cost (lesser of two evils) when confronted with two actions with comparable costs; (iii) the agent is subject to effects similar to cognitive dissonance and frustration. Neither of these features are shown by entropy maximization.

Comments:	21 pages, 3 figures
Subjects:	Data Analysis, Statistics and Probability (physics.data-an); Statistical Mechanics (cond-mat.stat-mech); Computer Science and Game Theory (cs.GT)
Cite as:	arXiv:1803.06638 [physics.data-an]
	(or arXiv:1803.06638v2 [physics.data-an] for this version)
	https://doi.org/10.48550/arXiv.1803.06638
Journal reference:	International Journal of Approximate Reasoning, 103, 270-287 (2018)

Submission history

From: Armen Allahverdyan [view email]
[v1] Sun, 18 Mar 2018 10:49:40 UTC (109 KB)
[v2] Sat, 1 Dec 2018 11:18:21 UTC (121 KB)

Physics > Data Analysis, Statistics and Probability

Title:Adaptive Decision Making via Entropy Minimization

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Physics > Data Analysis, Statistics and Probability

Title:Adaptive Decision Making via Entropy Minimization

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators