arXiv · 2110.08229
Influencing Towards Stable Multi-Agent Interactions
Abstract
Learning in multi-agent environments is difficult due to the non-stationarity introduced by an opponent's or partner's changing behaviors. Instead of reactively adapting to the other agent's (opponent or partner) behavior, we propose an algorithm to proactively influence the other agent's strategy to stabilize -- which can restrain the non-stationarity caused by the other agent. We learn a low-dimensional latent representation of the other agent's strategy and the dynamics of how the latent strategy evolves with respect to our robot's behavior. With this learned dynamics model, we can define an unsupervised stability reward to train our robot to deliberately influence the other agent to stabilize towards a single strategy. We demonstrate the effectiveness of stabilizing in improving efficiency of maximizing the task reward in a variety of simulated environments, including autonomous driving, emergent communication, and robotic manipulation. We show qualitative results on our website: https://sites.google.com/view/stable-marl/.
Explore related subjects
Keep this discovery
Woodrow Z. Wang, Andy Shih, Annie Xie, Dorsa Sadigh. 2021-10-05. Influencing Towards Stable Multi-Agent Interactions. https://arxiv.org/abs/2110.08229
Cite the original work for its findings. Save a collection to share your selection of sources.