TY - RPRT TI - Policy iteration for perfect information stochastic mean payoff games with bounded first return times is strongly polynomial AU - Marianne Akian AU - Stéphane Gaubert PY - 2013 UR - https://arxiv.org/abs/1310.4953 ID - 1310.4953 ER -