Polynomial time algorithm for optimal stopping with fixed accuracy

Goldberg, David A.; Chen, Yilun

Mathematics > Probability

arXiv:1807.02227 (math)

[Submitted on 6 Jul 2018 (v1), last revised 14 May 2024 (this version, v3)]

Title:Polynomial time algorithm for optimal stopping with fixed accuracy

Authors:David A. Goldberg, Yilun Chen

View PDF HTML (experimental)

Abstract:The problem of high-dimensional path-dependent optimal stopping (OS) is important to multiple academic communities and applications. Modern OS tasks often have a large number of decision epochs, and complicated non-Markovian dynamics, making them especially challenging. Standard approaches, often relying on ADP, duality, deep learning and other heuristics, have shown strong empirical performance, yet have limited rigorous guarantees (which may scale exponentially in the problem parameters and/or require previous knowledge of basis functions or additional continuity assumptions). Although past work has placed these problems in the framework of computational complexity and polynomial-time approximability, those analyses were limited to simple one-dimensional problems. For long-horizon complex OS problems, is a polynomial time solution even theoretically possible? We prove that given access to an efficient simulator of the underlying information process, and fixed accuracy epsilon, there exists an algorithm that returns an epsilon-optimal solution (both stopping policies and approximate optimal values) with computational complexity scaling polynomially in the time horizon and underlying dimension. Like the first polynomial-time (approximation) algorithms for several other well-studied problems, our theoretical guarantees are polynomial yet impractical. Our approach is based on a novel expansion for the optimal value which may be of independent interest.

Subjects:	Probability (math.PR); Data Structures and Algorithms (cs.DS); Optimization and Control (math.OC); Computational Finance (q-fin.CP); Mathematical Finance (q-fin.MF)
Cite as:	arXiv:1807.02227 [math.PR]
	(or arXiv:1807.02227v3 [math.PR] for this version)
	https://doi.org/10.48550/arXiv.1807.02227

Submission history

From: David Goldberg [view email]
[v1] Fri, 6 Jul 2018 02:53:45 UTC (69 KB)
[v2] Fri, 17 Aug 2018 18:47:02 UTC (77 KB)
[v3] Tue, 14 May 2024 23:24:31 UTC (71 KB)

Mathematics > Probability

Title:Polynomial time algorithm for optimal stopping with fixed accuracy

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Mathematics > Probability

Title:Polynomial time algorithm for optimal stopping with fixed accuracy

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators