TY - RPRT TI - An algorithm with nearly optimal pseudo-regret for both stochastic and adversarial bandits AU - Peter Auer AU - Chao-Kai Chiang PY - 2016 UR - https://arxiv.org/abs/1605.08722 ID - 1605.08722 ER -