TY - RPRT TI - Width-based Lookaheads with Learnt Base Policies and Heuristics Over the Atari-2600 Benchmark AU - Stefan O'Toole AU - Nir Lipovetzky AU - Miquel Ramirez AU - Adrian Pearce PY - 2021 UR - https://arxiv.org/abs/2106.12151 ID - 2106.12151 ER -