arXiv · 2210.06787
Observed Adversaries in Deep Reinforcement Learning
Abstract
In this work, we point out the problem of observed adversaries for deep policies. Specifically, recent work has shown that deep reinforcement learning is susceptible to adversarial attacks where an observed adversary acts under environmental constraints to invoke natural but adversarial observations. This setting is particularly relevant for HRI since HRI-related robots are expected to perform their tasks around and with other agents. In this work, we demonstrate that this effect persists even with low-dimensional observations. We further show that these adversarial attacks transfer across victims, which potentially allows malicious attackers to train an adversary without access to the target victim.
Explore related subjects
Keep this discovery
Eugene Lim, Harold Soh. 2022-10-13. Observed Adversaries in Deep Reinforcement Learning. https://arxiv.org/abs/2210.06787
Cite the original work for its findings. Save a collection to share your selection of sources.