Manipulating Reinforcement Learning: Poisoning Attacks on Cost Signals

TOP 文献データベース Manipulating Reinforcement Learning: Poisoning Attacks on Cost Signals

arxiv

AIセキュリティポータルbot

文献データベースの情報は、自動的に収集されています。

Source

https://arxiv.org/abs/2002.03827

PDF

https://arxiv.org/pdf/2002.03827

文献情報

作者: Yunhan Huang,Quanyan Zhu
公開日: 2020-2-8
更新日: 2020-7-21
所属機関: Department of Electrical and Computer Engineering, New York University
所属の国: United States of America
会議名

AIにより推定されたラベル

Q-Learningアルゴリズム敵対的攻撃収束分析

※ こちらのラベルはAIによって自動的に追加されました。そのため、正確でないことがあります。
詳細は文献データベースについてをご覧ください。

Abstract

This chapter studies emerging cyber-attacks on reinforcement learning (RL) and introduces a quantitative approach to analyze the vulnerabilities of RL. Focusing on adversarial manipulation on the cost signals, we analyze the performance degradation of TD($\lambda$) and $Q$-learning algorithms under the manipulation. For TD($\lambda$), the approximation learned from the manipulated costs has an approximation error bound proportional to the magnitude of the attack. The effect of the adversarial attacks on the bound does not depend on the choice of $\lambda$. In $Q$-learning, we show that $Q$-learning algorithms converge under stealthy attacks and bounded falsifications on cost signals. We characterize the relation between the falsified cost and the $Q$-factors as well as the policy learned by the learning agent which provides fundamental limits for feasible offensive and defensive moves. We propose a robust region in terms of the cost within which the adversary can never achieve the targeted policy. We provide conditions on the falsified cost which can mislead the agent to learn an adversary's favored policy. A case study of TD($\lambda$) learning is provided to corroborate the results.