TY - RPRT TI - Concave Utility Reinforcement Learning: the Mean-Field Game Viewpoint AU - Matthieu Geist AU - Julien Pérolat AU - Mathieu Laurière AU - Romuald Elie AU - Sarah Perrin AU - Olivier Bachem AU - Rémi Munos AU - Olivier Pietquin PY - 2022 UR - https://arxiv.org/abs/2106.03787 ID - 2106.03787 ER -