TY - RPRT TI - Variance-Based Rewards for Approximate Bayesian Reinforcement Learning AU - Jonathan Sorg AU - Satinder Singh AU - Richard L. Lewis PY - 2012 UR - https://arxiv.org/abs/1203.3518 ID - 1203.3518 ER -