TY - RPRT TI - Reinforcement Learning with Linear Function Approximation and LQ control Converges AU - Istvan Szita AU - Andras Lorincz PY - 2007 UR - https://arxiv.org/abs/cs/0306120 ID - cs/0306120 ER -