TY - RPRT TI - Fast Reinforcement Learning with Large Action Sets using Error-Correcting Output Codes for MDP Factorization AU - Gabriel Dulac-Arnold AU - Ludovic Denoyer AU - Philippe Preux AU - Patrick Gallinari PY - 2012 UR - https://arxiv.org/abs/1203.0203 ID - 1203.0203 ER -