TY - RPRT TI - An Online Prediction Algorithm for Reinforcement Learning with Linear Function Approximation using Cross Entropy Method AU - Ajin George Joseph AU - Shalabh Bhatnagar PY - 2018 UR - https://arxiv.org/abs/1806.06720 ID - 1806.06720 ER -