TY - RPRT TI - Deterministic limit of temporal difference reinforcement learning for stochastic games AU - Wolfram Barfuss AU - Jonathan F. Donges AU - Jürgen Kurths PY - 2019 DO - 10.1103/physreve.99.043305 UR - https://arxiv.org/abs/1809.07225 ID - 1809.07225 ER -