TY - RPRT TI - Optimal Regret Bounds for Selecting the State Representation in Reinforcement Learning AU - Odalric-Ambrym Maillard AU - Phuong Nguyen AU - Ronald Ortner AU - Daniil Ryabko PY - 2013 UR - https://arxiv.org/abs/1302.2553 ID - 1302.2553 ER -