TY - RPRT TI - Optimal Learning for Sequential Decision Making for Expensive Cost Functions with Stochastic Binary Feedbacks AU - Yingfei Wang AU - Chu Wang AU - Warren Powell PY - 2017 UR - https://arxiv.org/abs/1709.05216 ID - 1709.05216 ER -