TY - RPRT TI - Reward Prediction Error as an Exploration Objective in Deep RL AU - Riley Simmons-Edler AU - Ben Eisner AU - Daniel Yang AU - Anthony Bisulco AU - Eric Mitchell AU - Sebastian Seung AU - Daniel Lee PY - 2021 UR - https://arxiv.org/abs/1906.08189 ID - 1906.08189 ER -