Contrastive Modules with Temporal Attention for Multi-Task Reinforcement Learning

Lan, Siming; Zhang, Rui; Yi, Qi; Guo, Jiaming; Peng, Shaohui; Gao, Yunkai; Wu, Fan; Chen, Ruizhi; Du, Zidong; Hu, Xing; Zhang, Xishan; Li, Ling; Chen, Yunji

Computer Science > Machine Learning

arXiv:2311.01075 (cs)

[Submitted on 2 Nov 2023]

Title:Contrastive Modules with Temporal Attention for Multi-Task Reinforcement Learning

Authors:Siming Lan, Rui Zhang, Qi Yi, Jiaming Guo, Shaohui Peng, Yunkai Gao, Fan Wu, Ruizhi Chen, Zidong Du, Xing Hu, Xishan Zhang, Ling Li, Yunji Chen

View PDF

Abstract:In the field of multi-task reinforcement learning, the modular principle, which involves specializing functionalities into different modules and combining them appropriately, has been widely adopted as a promising approach to prevent the negative transfer problem that performance degradation due to conflicts between tasks. However, most of the existing multi-task RL methods only combine shared modules at the task level, ignoring that there may be conflicts within the task. In addition, these methods do not take into account that without constraints, some modules may learn similar functions, resulting in restricting the model's expressiveness and generalization capability of modular methods. In this paper, we propose the Contrastive Modules with Temporal Attention(CMTA) method to address these limitations. CMTA constrains the modules to be different from each other by contrastive learning and combining shared modules at a finer granularity than the task level with temporal attention, alleviating the negative transfer within the task and improving the generalization ability and the performance for multi-task RL. We conducted the experiment on Meta-World, a multi-task RL benchmark containing various robotics manipulation tasks. Experimental results show that CMTA outperforms learning each task individually for the first time and achieves substantial performance improvements over the baselines.

Comments:	This paper has been accepted at NeurIPS 2023 as a poster
Subjects:	Machine Learning (cs.LG)
Cite as:	arXiv:2311.01075 [cs.LG]
	(or arXiv:2311.01075v1 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2311.01075

Submission history

From: Siming Lan [view email]
[v1] Thu, 2 Nov 2023 08:41:00 UTC (7,722 KB)

✅2024-10-01: arxiv.org is back to normal.✅

Computer Science > Machine Learning

Title:Contrastive Modules with Temporal Attention for Multi-Task Reinforcement Learning

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

✅2024-10-01: arxiv.org is back to normal.✅

Computer Science > Machine Learning

Title:Contrastive Modules with Temporal Attention for Multi-Task Reinforcement Learning

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators