arXiv · 1310.7610
Distributed Reinforcement Learning via Gossip
Abstract
We consider the classical TD(0) algorithm implemented on a network of agents wherein the agents also incorporate the updates received from neighboring agents using a gossip-like mechanism. The combined scheme is shown to converge for both discounted and average cost problems.
Explore related subjects
Keep this discovery
Adwaitvedant S. Mathkar, Vivek S. Borkar. 2013-10-28. Distributed Reinforcement Learning via Gossip. https://arxiv.org/abs/1310.7610
Cite the original work for its findings. Save a collection to share your selection of sources.