TY - RPRT TI - A neural network based policy iteration algorithm with global $H^2$-superlinear convergence for stochastic games on domains AU - Kazufumi Ito AU - Christoph Reisinger AU - Yufei Zhang PY - 2020 UR - https://arxiv.org/abs/1906.02304 ID - 1906.02304 ER -