Dynamics Generalization via Information Bottleneck in Deep Reinforcement Learning

Lu, Xingyu; Lee, Kimin; Abbeel, Pieter; Tiomkin, Stas

Computer Science > Machine Learning

arXiv:2008.00614 (cs)

[Submitted on 3 Aug 2020]

Title:Dynamics Generalization via Information Bottleneck in Deep Reinforcement Learning

Authors:Xingyu Lu, Kimin Lee, Pieter Abbeel, Stas Tiomkin

View PDF

Abstract:Despite the significant progress of deep reinforcement learning (RL) in solving sequential decision making problems, RL agents often overfit to training environments and struggle to adapt to new, unseen environments. This prevents robust applications of RL in real world situations, where system dynamics may deviate wildly from the training settings. In this work, our primary contribution is to propose an information theoretic regularization objective and an annealing-based optimization method to achieve better generalization ability in RL agents. We demonstrate the extreme generalization benefits of our approach in different domains ranging from maze navigation to robotic tasks; for the first time, we show that agents can generalize to test parameters more than 10 standard deviations away from the training parameter distribution. This work provides a principled way to improve generalization in RL by gradually removing information that is redundant for task-solving; it opens doors for the systematic study of generalization from training to extremely different testing settings, focusing on the established connections between information theory and machine learning.

Comments:	16 pages
Subjects:	Machine Learning (cs.LG); Artificial Intelligence (cs.AI); Machine Learning (stat.ML)
Cite as:	arXiv:2008.00614 [cs.LG]
	(or arXiv:2008.00614v1 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2008.00614

Submission history

From: Xingyu Lu [view email]
[v1] Mon, 3 Aug 2020 02:24:20 UTC (7,190 KB)

Full-text links:

Access Paper:

view license

Current browse context:

cs.LG

< prev | next >

new | recent | 2020-08

Change to browse by:

cs
cs.AI
stat
stat.ML

References & Citations

DBLP - CS Bibliography

listing | bibtex

Kimin Lee
Pieter Abbeel
Stas Tiomkin

export BibTeX citation

Computer Science > Machine Learning

Title:Dynamics Generalization via Information Bottleneck in Deep Reinforcement Learning

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Dynamics Generalization via Information Bottleneck in Deep Reinforcement Learning

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators