TY - RPRT TI - Adaptive Temporal-Difference Learning for Policy Evaluation with Per-State Uncertainty Estimates AU - Hugo Penedones AU - Carlos Riquelme AU - Damien Vincent AU - Hartmut Maennel AU - Timothy Mann AU - Andre Barreto AU - Sylvain Gelly AU - Gergely Neu PY - 2019 UR - https://arxiv.org/abs/1906.07987 ID - 1906.07987 ER -