TY - RPRT TI - Fast and Data Efficient Reinforcement Learning from Pixels via Non-Parametric Value Approximation AU - Alexander Long AU - Alan Blair AU - Herke van Hoof PY - 2022 UR - https://arxiv.org/abs/2203.03078 ID - 2203.03078 ER -