TY - RPRT TI - Hierarchical Reinforcement Learning with the MAXQ Value Function Decomposition AU - Thomas G. Dietterich PY - 1999 UR - https://arxiv.org/abs/cs/9905014 ID - cs/9905014 ER -