TY - RPRT TI - Reinforcement Learning with Wasserstein Distance Regularisation, with Applications to Multipolicy Learning AU - Mohammed Amin Abdullah AU - Aldo Pacchiano AU - Moez Draief PY - 2019 UR - https://arxiv.org/abs/1802.03976 ID - 1802.03976 ER -