TY - RPRT TI - Training Reinforcement Neurocontrollers Using the Polytope Algorithm AU - A. Likas AU - I. E. Lagaris PY - 1998 UR - https://arxiv.org/abs/cs/9812002 ID - cs/9812002 ER -