TY - RPRT TI - DQN with model-based exploration: efficient learning on environments with sparse rewards AU - Stephen Zhen Gou AU - Yuyang Liu PY - 2019 UR - https://arxiv.org/abs/1903.09295 ID - 1903.09295 ER -