Reinforcement Learning with Temporal-Logic-Based Causal Diagrams

Paliwal, Yash; Roy, Rajarshi; Gaglione, Jean-Raphaël; Baharisangari, Nasim; Neider, Daniel; Duan, Xiaoming; Topcu, Ufuk; Xu, Zhe

Computer Science > Artificial Intelligence

arXiv:2306.13732 (cs)

[Submitted on 23 Jun 2023]

Title:Reinforcement Learning with Temporal-Logic-Based Causal Diagrams

Authors:Yash Paliwal, Rajarshi Roy, Jean-Raphaël Gaglione, Nasim Baharisangari, Daniel Neider, Xiaoming Duan, Ufuk Topcu, Zhe Xu

View PDF

Abstract:We study a class of reinforcement learning (RL) tasks where the objective of the agent is to accomplish temporally extended goals. In this setting, a common approach is to represent the tasks as deterministic finite automata (DFA) and integrate them into the state-space for RL algorithms. However, while these machines model the reward function, they often overlook the causal knowledge about the environment. To address this limitation, we propose the Temporal-Logic-based Causal Diagram (TL-CD) in RL, which captures the temporal causal relationships between different properties of the environment. We exploit the TL-CD to devise an RL algorithm in which an agent requires significantly less exploration of the environment. To this end, based on a TL-CD and a task DFA, we identify configurations where the agent can determine the expected rewards early during an exploration. Through a series of case studies, we demonstrate the benefits of using TL-CDs, particularly the faster convergence of the algorithm to an optimal policy due to reduced exploration of the environment.

Subjects:	Artificial Intelligence (cs.AI); Formal Languages and Automata Theory (cs.FL)
Cite as:	arXiv:2306.13732 [cs.AI]
	(or arXiv:2306.13732v1 [cs.AI] for this version)
	https://doi.org/10.48550/arXiv.2306.13732

Submission history

From: Jean-Raphaël Gaglione [view email]
[v1] Fri, 23 Jun 2023 18:42:27 UTC (1,542 KB)

Computer Science > Artificial Intelligence

Title:Reinforcement Learning with Temporal-Logic-Based Causal Diagrams

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Artificial Intelligence

Title:Reinforcement Learning with Temporal-Logic-Based Causal Diagrams

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators