TY - RPRT TI - Agent-Agnostic Human-in-the-Loop Reinforcement Learning AU - David Abel AU - John Salvatier AU - Andreas Stuhlmüller AU - Owain Evans PY - 2017 UR - https://arxiv.org/abs/1701.04079 ID - 1701.04079 ER -