Solving Offline Reinforcement Learning with Decision Tree Regression

Koirala, Prajwal; Fleming, Cody

Computer Science > Machine Learning

arXiv:2401.11630 (cs)

[Submitted on 21 Jan 2024 (v1), last revised 14 Oct 2024 (this version, v2)]

Title:Solving Offline Reinforcement Learning with Decision Tree Regression

Authors:Prajwal Koirala, Cody Fleming

View PDF HTML (experimental)

Abstract:This study presents a novel approach to addressing offline reinforcement learning (RL) problems by reframing them as regression tasks that can be effectively solved using Decision Trees. Mainly, we introduce two distinct frameworks: return-conditioned and return-weighted decision tree policies (RCDTP and RWDTP), both of which achieve notable speed in agent training as well as inference, with training typically lasting less than a few minutes. Despite the simplification inherent in this reformulated approach to offline RL, our agents demonstrate performance that is at least on par with the established methods. We evaluate our methods on D4RL datasets for locomotion and manipulation, as well as other robotic tasks involving wheeled and flying robots. Additionally, we assess performance in delayed/sparse reward scenarios and highlight the explainability of these policies through action distribution and feature importance.

Subjects:	Machine Learning (cs.LG); Systems and Control (eess.SY)
Cite as:	arXiv:2401.11630 [cs.LG]
	(or arXiv:2401.11630v2 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2401.11630

Submission history

From: Prajwal Koirala [view email]
[v1] Sun, 21 Jan 2024 23:50:46 UTC (6,378 KB)
[v2] Mon, 14 Oct 2024 22:13:31 UTC (1,053 KB)

Full-text links:

Access Paper:

view license

Current browse context:

cs.LG

< prev | next >

new | recent | 2024-01

Change to browse by:

cs
cs.SY
eess
eess.SY

References & Citations

export BibTeX citation

Computer Science > Machine Learning

Title:Solving Offline Reinforcement Learning with Decision Tree Regression

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Solving Offline Reinforcement Learning with Decision Tree Regression

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators