Asynchronous Coagent Networks

Kostas, James E.; Nota, Chris; Thomas, Philip S.

Computer Science > Machine Learning

arXiv:1902.05650 (cs)

[Submitted on 15 Feb 2019 (v1), last revised 10 Aug 2020 (this version, v4)]

Title:Asynchronous Coagent Networks

Authors:James E. Kostas, Chris Nota, Philip S. Thomas

View PDF

Abstract:Coagent policy gradient algorithms (CPGAs) are reinforcement learning algorithms for training a class of stochastic neural networks called coagent networks. In this work, we prove that CPGAs converge to locally optimal policies. Additionally, we extend prior theory to encompass asynchronous and recurrent coagent networks. These extensions facilitate the straightforward design and analysis of hierarchical reinforcement learning algorithms like the option-critic, and eliminate the need for complex derivations of customized learning rules for these algorithms.

Comments:	Updated version
Subjects:	Machine Learning (cs.LG); Machine Learning (stat.ML)
Cite as:	arXiv:1902.05650 [cs.LG]
	(or arXiv:1902.05650v4 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.1902.05650

Submission history

From: James Kostas [view email]
[v1] Fri, 15 Feb 2019 00:16:10 UTC (779 KB)
[v2] Mon, 18 Feb 2019 16:27:13 UTC (779 KB)
[v3] Thu, 21 Feb 2019 22:31:58 UTC (779 KB)
[v4] Mon, 10 Aug 2020 04:57:55 UTC (810 KB)

Full-text links:

Access Paper:

view license

Current browse context:

stat

< prev | next >

new | recent | 2019-02

Change to browse by:

cs
cs.LG
stat.ML

References & Citations

DBLP - CS Bibliography

listing | bibtex

James Kostas
Chris Nota
Philip S. Thomas

export BibTeX citation

Computer Science > Machine Learning

Title:Asynchronous Coagent Networks

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Asynchronous Coagent Networks

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators