Statistical Analysis of Policy Space Compression Problem

Molaei, Majid; Restelli, Marcello; Metelli, Alberto Maria; Papini, Matteo

Computer Science > Machine Learning

arXiv:2411.09900 (cs)

[Submitted on 15 Nov 2024]

Title:Statistical Analysis of Policy Space Compression Problem

Authors:Majid Molaei, Marcello Restelli, Alberto Maria Metelli, Matteo Papini

View PDF HTML (experimental)

Abstract:Policy search methods are crucial in reinforcement learning, offering a framework to address continuous state-action and partially observable problems. However, the complexity of exploring vast policy spaces can lead to significant inefficiencies. Reducing the policy space through policy compression emerges as a powerful, reward-free approach to accelerate the learning process. This technique condenses the policy space into a smaller, representative set while maintaining most of the original effectiveness. Our research focuses on determining the necessary sample size to learn this compressed set accurately. We employ Rényi divergence to measure the similarity between true and estimated policy distributions, establishing error bounds for good approximations. To simplify the analysis, we employ the $l_1$ norm, determining sample size requirements for both model-based and model-free settings. Finally, we correlate the error bounds from the $l_1$ norm with those from Rényi divergence, distinguishing between policies near the vertices and those in the middle of the policy space, to determine the lower and upper bounds for the required sample sizes.

Subjects:	Machine Learning (cs.LG); Artificial Intelligence (cs.AI); Machine Learning (stat.ML)
Cite as:	arXiv:2411.09900 [cs.LG]
	(or arXiv:2411.09900v1 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2411.09900

Submission history

From: Majid Molaei [view email]
[v1] Fri, 15 Nov 2024 02:46:55 UTC (11 KB)

Computer Science > Machine Learning

Title:Statistical Analysis of Policy Space Compression Problem

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Statistical Analysis of Policy Space Compression Problem

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators