TY - RPRT TI - Truncating Temporal Differences: On the Efficient Implementation of TD(lambda) for Reinforcement Learning AU - P. Cichosz PY - 1995 UR - https://arxiv.org/abs/cs/9501103 ID - cs/9501103 ER -