TY - RPRT TI - Instance-Dependent Complexity of Contextual Bandits and Reinforcement Learning: A Disagreement-Based Perspective AU - Dylan J. Foster AU - Alexander Rakhlin AU - David Simchi-Levi AU - Yunzong Xu PY - 2020 UR - https://arxiv.org/abs/2010.03104 ID - 2010.03104 ER -