TY - RPRT TI - Posterior Sampling for Large Scale Reinforcement Learning AU - Georgios Theocharous AU - Zheng Wen AU - Yasin Abbasi-Yadkori AU - Nikos Vlassis PY - 2018 UR - https://arxiv.org/abs/1711.07979 ID - 1711.07979 ER -