arXiv · 1907.05689
Gittins' theorem under uncertainty
Abstract
We study dynamic allocation problems for discrete time multi-armed bandits under uncertainty, based on the the theory of nonlinear expectations. We show that, under strong independence of the bandits and with some relaxation in the definition of optimality, a Gittins allocation index gives optimal choices. This involves studying the interaction of our uncertainty with controls which determine the filtration. We also run a simple numerical example which illustrates the interaction between the willingness to explore and uncertainty aversion of the agent when making decisions.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Samuel N. Cohen, Tanut Treetanthiploet. 2019-07-12. Gittins' theorem under uncertainty. https://arxiv.org/abs/1907.05689
Cite the original work for its findings. Save a collection to share your selection of sources.