Reinforcement Learning Without Backpropagation or a Clock

Kostas, James; Nota, Chris; Thomas, Philip S.

Computer Science > Machine Learning

arXiv:1902.05650v1 (cs)

[Submitted on 15 Feb 2019 (this version), latest version 10 Aug 2020 (v4)]

Title:Reinforcement Learning Without Backpropagation or a Clock

Authors:James Kostas, Chris Nota, Philip S. Thomas

View PDF

Abstract:In this paper we introduce a reinforcement learning (RL) approach for training policies, including artificial neural network policies, that is both \emph{backpropagation-free} and \emph{clock-free}. It is \emph{backpropagation-free} in that it does not propagate any information backwards through the network. It is \emph{clock-free} in that no signal is given to each node in the network to specify when it should compute its output and when it should update its weights. We contend that these two properties increase the biological plausibility of our algorithms and facilitate distributed implementations. Additionally, our approach eliminates the need for customized learning rules for hierarchical RL algorithms like the option-critic.

Subjects:	Machine Learning (cs.LG); Machine Learning (stat.ML)
Cite as:	arXiv:1902.05650 [cs.LG]
	(or arXiv:1902.05650v1 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.1902.05650

Submission history

From: James Kostas [view email]
[v1] Fri, 15 Feb 2019 00:16:10 UTC (779 KB)
[v2] Mon, 18 Feb 2019 16:27:13 UTC (779 KB)
[v3] Thu, 21 Feb 2019 22:31:58 UTC (779 KB)
[v4] Mon, 10 Aug 2020 04:57:55 UTC (810 KB)

Full-text links:

Access Paper:

view license

Current browse context:

cs.LG

< prev | next >

new | recent | 2019-02

Change to browse by:

cs
stat
stat.ML

References & Citations

DBLP - CS Bibliography

listing | bibtex

James Kostas
Chris Nota
Philip S. Thomas

export BibTeX citation

Computer Science > Machine Learning

Title:Reinforcement Learning Without Backpropagation or a Clock

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Reinforcement Learning Without Backpropagation or a Clock

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators