TY - RPRT TI - Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm AU - David Silver AU - Thomas Hubert AU - Julian Schrittwieser AU - Ioannis Antonoglou AU - Matthew Lai AU - Arthur Guez AU - Marc Lanctot AU - Laurent Sifre AU - Dharshan Kumaran AU - Thore Graepel AU - Timothy Lillicrap AU - Karen Simonyan AU - Demis Hassabis PY - 2017 UR - https://arxiv.org/abs/1712.01815 ID - 1712.01815 ER -