Continuous-time mean field Markov decision models

Bäuerle, Nicole; Höfer, Sebastian

Mathematics > Probability

arXiv:2307.01575 (math)

[Submitted on 4 Jul 2023 (v1), last revised 12 Jun 2024 (this version, v3)]

Title:Continuous-time mean field Markov decision models

Authors:Nicole Bäuerle, Sebastian Höfer

View PDF HTML (experimental)

Abstract:We consider a finite number of $N$ statistically equal agents, each moving on a finite set of states according to a continuous-time Markov Decision Process (MDP). Transition intensities of the agents and generated rewards depend not only on the state and action of the agent itself, but also on the states of the other agents as well as the chosen action. Interactions like this are typical for a wide range of models in e.g. biology, epidemics, finance, social science and queueing systems among others. The aim is to maximize the expected discounted reward of the system, i.e. the agents have to cooperate as a team. Computationally this is a difficult task when $N$ is large. Thus, we consider the limit for $N\to\infty.$ In contrast to other papers we treat this problem from an MDP perspective. This has the advantage that we need less regularity assumptions in order to construct asymptotically optimal strategies than using viscosity solutions of HJB equations. The convergence rate is $1/\sqrt{N}$. We show how to apply our results using two examples: a machine replacement problem and a problem from epidemics. We also show that optimal feedback policies from the limiting problem are not necessarily asymptotically optimal.

Subjects:	Probability (math.PR); Optimization and Control (math.OC)
MSC classes:	90C40, 60J27
Cite as:	arXiv:2307.01575 [math.PR]
	(or arXiv:2307.01575v3 [math.PR] for this version)
	https://doi.org/10.48550/arXiv.2307.01575

Submission history

From: Nicole Bäuerle [view email]
[v1] Tue, 4 Jul 2023 09:04:36 UTC (507 KB)
[v2] Sun, 12 Nov 2023 16:04:40 UTC (506 KB)
[v3] Wed, 12 Jun 2024 08:59:52 UTC (601 KB)

Mathematics > Probability

Title:Continuous-time mean field Markov decision models

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Mathematics > Probability

Title:Continuous-time mean field Markov decision models

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators