TY - RPRT TI - Deep Learning for Reward Design to Improve Monte Carlo Tree Search in ATARI Games AU - Xiaoxiao Guo AU - Satinder Singh AU - Richard Lewis AU - Honglak Lee PY - 2016 UR - https://arxiv.org/abs/1604.07095 ID - 1604.07095 ER -