Deceptive Reinforcement Learning Under Adversarial Manipulations on Cost Signals

Huang, Yunhan; Zhu, Quanyan

Computer Science > Machine Learning

arXiv:1906.10571 (cs)

[Submitted on 24 Jun 2019 (v1), last revised 16 Aug 2019 (this version, v3)]

Title:Deceptive Reinforcement Learning Under Adversarial Manipulations on Cost Signals

Authors:Yunhan Huang, Quanyan Zhu

View PDF

Abstract:This paper studies reinforcement learning (RL) under malicious falsification on cost signals and introduces a quantitative framework of attack models to understand the vulnerabilities of RL. Focusing on $Q$-learning, we show that $Q$-learning algorithms converge under stealthy attacks and bounded falsifications on cost signals. We characterize the relation between the falsified cost and the $Q$-factors as well as the policy learned by the learning agent which provides fundamental limits for feasible offensive and defensive moves. We propose a robust region in terms of the cost within which the adversary can never achieve the targeted policy. We provide conditions on the falsified cost which can mislead the agent to learn an adversary's favored policy. A numerical case study of water reservoir control is provided to show the potential hazards of RL in learning-based control systems and corroborate the results.

Subjects:	Machine Learning (cs.LG); Artificial Intelligence (cs.AI); Optimization and Control (math.OC)
Cite as:	arXiv:1906.10571 [cs.LG]
	(or arXiv:1906.10571v3 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.1906.10571

Submission history

From: Yunhan Huang [view email]
[v1] Mon, 24 Jun 2019 15:48:54 UTC (3,096 KB)
[v2] Tue, 13 Aug 2019 14:55:39 UTC (3,101 KB)
[v3] Fri, 16 Aug 2019 18:32:49 UTC (3,099 KB)

Full-text links:

Access Paper:

view license

Current browse context:

cs.LG

< prev | next >

new | recent | 2019-06

Change to browse by:

cs
cs.AI
math
math.OC

References & Citations

DBLP - CS Bibliography

listing | bibtex

Yunhan Huang
Quanyan Zhu

export BibTeX citation

Computer Science > Machine Learning

Title:Deceptive Reinforcement Learning Under Adversarial Manipulations on Cost Signals

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Deceptive Reinforcement Learning Under Adversarial Manipulations on Cost Signals

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators