TY - RPRT TI - Tightening Exploration in Upper Confidence Reinforcement Learning AU - Hippolyte Bourel AU - Odalric-Ambrym Maillard AU - Mohammad Sadegh Talebi PY - 2021 UR - https://arxiv.org/abs/2004.09656 ID - 2004.09656 ER -