TY - RPRT TI - Distributed Bandit Learning: Near-Optimal Regret with Efficient Communication AU - Yuanhao Wang AU - Jiachen Hu AU - Xiaoyu Chen AU - Liwei Wang PY - 2019 UR - https://arxiv.org/abs/1904.06309 ID - 1904.06309 ER -