Modular Adaptive Policy Selection for Multi-Task Imitation Learning through Task Division

Antotsiou, Dafni; Ciliberto, Carlo; Kim, Tae-Kyun

Computer Science > Machine Learning

arXiv:2203.14855 (cs)

[Submitted on 28 Mar 2022 (v1), last revised 13 May 2022 (this version, v2)]

Title:Modular Adaptive Policy Selection for Multi-Task Imitation Learning through Task Division

Authors:Dafni Antotsiou, Carlo Ciliberto, Tae-Kyun Kim

View PDF

Abstract:Deep imitation learning requires many expert demonstrations, which can be hard to obtain, especially when many tasks are involved. However, different tasks often share similarities, so learning them jointly can greatly benefit them and alleviate the need for many demonstrations. But, joint multi-task learning often suffers from negative transfer, sharing information that should be task-specific. In this work, we introduce a method to perform multi-task imitation while allowing for task-specific features. This is done by using proto-policies as modules to divide the tasks into simple sub-behaviours that can be shared. The proto-policies operate in parallel and are adaptively chosen by a selector mechanism that is jointly trained with the modules. Experiments on different sets of tasks show that our method improves upon the accuracy of single agents, task-conditioned and multi-headed multi-task agents, as well as state-of-the-art meta learning agents. We also demonstrate its ability to autonomously divide the tasks into both shared and task-specific sub-behaviours.

Comments:	ICRA 2022 contribution paper
Subjects:	Machine Learning (cs.LG); Robotics (cs.RO)
Cite as:	arXiv:2203.14855 [cs.LG]
	(or arXiv:2203.14855v2 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2203.14855

Submission history

From: Dafni Antotsiou M.Sc. [view email]
[v1] Mon, 28 Mar 2022 15:53:17 UTC (3,771 KB)
[v2] Fri, 13 May 2022 10:48:38 UTC (1,898 KB)

Computer Science > Machine Learning

Title:Modular Adaptive Policy Selection for Multi-Task Imitation Learning through Task Division

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Modular Adaptive Policy Selection for Multi-Task Imitation Learning through Task Division

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators